mlx-lm
★ 7.2KMLX LM is a Python package for running and fine-tuning LLMs on Apple silicon via MLX. It offers CLI and Python APIs for generation, chat, LoRA/full fine-tuning, quantization, GGUF conversion, prompt caching, and an OpenAI-compatible server.
AI Frameworks | Python · llm-inference · apple-silicon
View Project →