What is llama.cpp?
llama.cpp is an open-source C/C++ inference engine for running large language models locally on CPUs and consumer GPUs, using the GGUF format.
Primary sources
Specifications cited by the tools below whose own titles define llama.cpp.
2 tools that work with llama.cpp
Files never leave your browser.
Related terms
Terms that appear alongside llama.cpp on the same tools.