Vibe Coding Discover

Use Cases

Serve A Local Model Through Openai-compatible Apis

Published projects tagged with this use case.

1 project

Strata

★ 8.1K

Strata runs Qwen3.8-Flash-Next on consumer NVIDIA or AMD hardware, using GPU, RAM, and SSD resources to fit the large model. It provides a browser chat UI and local OpenAI-compatible, Anthropic-compatible, and MCP interfaces.

AI Frameworks | C++ · local inference · LLM serving

View Project →