llama.cpp
★ 129KHigh-performance C/C++ inference for local LLMs across CPUs, GPUs, and devices.
AI Frameworks | C++ · llama.cpp · Inference
View Project →Published projects tagged with this use case.
1 project
High-performance C/C++ inference for local LLMs across CPUs, GPUs, and devices.
AI Frameworks | C++ · llama.cpp · Inference
View Project →