Vibe Coding Discover

Use Cases

Serve An Openai-compatible Local Inference Server

Published projects tagged with this use case.

1 project

mlx-lm

★ 7.2K

MLX LM is a Python package for running and fine-tuning LLMs on Apple silicon via MLX. It offers CLI and Python APIs for generation, chat, LoRA/full fine-tuning, quantization, GGUF conversion, prompt caching, and an OpenAI-compatible server.

AI Frameworks | Python · llm-inference · apple-silicon

View Project →