Running Large MoE LLMs on Modest Hardware with FreeToken
A new open-source inference engine designed to run MoE models on hardware that doesn't have enough VRAM to hold them
Sep 10, 202615 min read14

Search for a command to run...
Articles tagged with #local-llm
A new open-source inference engine designed to run MoE models on hardware that doesn't have enough VRAM to hold them

or... "necessity is the mother of invention"
