llama.cpp
View on GitHubLLM inference in C/C++.
High-performance C/C++ inference for local LLMs across CPUs, GPUs, and devices.
Use Cases
Local LLMEdge InferencePrivate Apps
Built With
- Language
- C++
- Frameworks
- llama.cpp
Tags
C++ · llama.cpp · Inference