Vibe Coding Discover

Use Cases

Serve Local Models Through Openai-compatible Apis

Published projects tagged with this use case.

1 project

FreeToken

★ 14K

FreeToken is a local MoE inference and serving engine that runs large open-weight models across GPU, CPU, and host memory. It provides OpenAI- and Anthropic-compatible APIs, plus a desktop app and CLI.

AI Frameworks | Python · LLM inference · MoE serving

View Project →