

llama.cpp
#3 en Runtimes LLM locauxGeorgi Gerganov · depuis March 10, 2023 · 3× · vu le 09 oct. 2026
75
Momentum
llama.cpp is an open-source C/C++ library for LLM inference, originally developed by Georgi Gerganov and now maintained within the ggml-org project. It runs locally on CPU and GPU and forms the basis of many local LLM tools such as Ollama and LM Studio. The official website llama.app provides an installer and getting-started documentation.
Historique du momentum
11.07.09.10.
Fonctionnalités
| License | MIT |
| Platform | Windows, macOS, Linux (Docker, Homebrew) |
| Price | Free (open source) |
| Protocol Compatibility | OpenAI-compatible API (/v1/chat/completions, /v1/embeddings), Anthropic Messages-compatible, Ollama shim (/api/tags) |
| Release Date | March 10, 2023 |
| Supported Models/Providers | GGUF models (quantized), e.g. Qwen, Llama, Mistral, Gemma |