laya-mlx
★ 5.8KNative Apple Silicon MLX inference runtime for Laya typed-decision models. Returns choice probabilities, rubric scores and P(true) locally in ~7-14 ms per short question, with no PyTorch, tokenizer decoding or cloud API at runtime.
AI Frameworks | Python · mlx · apple-silicon
View Project →