Vibe Coding Discover

Use Cases

Run Large Moe Models Locally On Consumer Hardware

Published projects tagged with this use case.

1 project

FreeToken

★ 14K

FreeToken is a local MoE inference and serving engine that runs large open-weight models across GPU, CPU, and host memory. It provides OpenAI- and Anthropic-compatible APIs, plus a desktop app and CLI.

AI Frameworks | Python · LLM inference · MoE serving

View Project →