LLM Perks

llama.cpp

ggml-org · Local LLM runners
Open sourceBYOKOpenAI API★ 129K

Reference C/C++ engine for GGUF models; llama-server provides an OpenAI-compatible API and web UI on CPU, CUDA, Metal, Vulkan and ROCm.

License
MIT
Platforms
Windows, macOS, Linux, Docker, CLI
Own API keys
Yes
Checked
2026-09-24
Connect a free model

llama.cpp works with OpenAI-compatible APIs. Grab a free key and a model from the catalog: for coding · for agents · providers and base URLs

llama.cpp alternatives

FAQ

Is llama.cpp free?

Completely free and open source, with no accounts, API keys or usage limits.

Does llama.cpp work with free models?

Yes. It supports your own keys and OpenAI-compatible APIs, so you can plug in free models from OpenRouter, Groq, NVIDIA Build, or a local Ollama.

What license does llama.cpp use?

MIT. The source code is open.