Skip to content
Convertto

What is llama.cpp?

llama.cpp is an open-source C/C++ inference engine for running large language models locally on CPUs and consumer GPUs, using the GGUF format.

Primary sources

Specifications cited by the tools below whose own titles define llama.cpp.

2 tools that work with llama.cpp

Files never leave your browser.

Related terms

Terms that appear alongside llama.cpp on the same tools.