comparar
llama.cpp vs Ollama
Los mismos datos para ambos, leídos de GitHub cada noche, y la relación que revisó una persona.
| Dato | llama.cpp | Ollama |
|---|---|---|
| Lenguaje | C++ | Go |
| Licencia | MIT | MIT |
| Estrellas | 130k | 182k |
| Última versión | v0.6.0 | v0.35.1 |
| Último push | 2026-10-06 | 2026-10-06 |
| Ritmo de versiones | muy pocas versiones | alrededor de 4 días entre versiones |
| Colaboradores activos | 173+ autores de commits en la rama por defecto en los últimos 90 días | 18 autores de commits en la rama por defecto en los últimos 90 días |
| Avisos | ninguno | ninguno |
Cómo se relacionan
Ambos sustituyen a ChatGPT. Alternativas a ChatGPT →
Parcialllama.cppRuns open-weight models on CPU or GPU, with a built-in server and a minimal web chat.
ParcialOllamaRuns open-weight models on your own machine behind a local API. Pair it with a chat app for the interface.
Ambos sustituyen a Claude. Alternativas a Claude →
Parcialllama.cppRuns open-weight models on CPU or GPU, with a built-in server and a minimal web chat.
ParcialOllamaRuns open-weight models on your own machine behind a local API. Pair it with a chat app for the interface.
llama.cpp
- b114452026-10-06vulkan : check for null vkEnumerateInstanceVersion (#29872)versión preliminar
- b114432026-10-06models : consolidate nextn row cropping into shared helpers (#30017)versión preliminar
- b114402026-10-06llama : re-reserve the sched when the nextn extraction flags change (#30020)versión preliminar
- b114392026-10-06ggml: refactor selective expert copying to user code (#29943)versión preliminar
- b114382026-10-06test-llama-archs : initialize backends before generating models (#30034)versión preliminar
Ollama
- v0.40.0-rc62026-09-25v0.40.0versión preliminar
- v0.35.12026-09-29Ollama now supports Clef and Clef Flash, Cloudflare's new open-source decision models, through /v1/systemone.
- v0.35.02026-09-28Ollama now supports decision models through /v1/systemone, based on TypeSafe’s Jev API.
- v0.34.42026-09-23Structured outputs on thinking models now apply in a single pass, making them faster and more reliable.
- v0.34.32026-09-19GET /api/show now advertises each model's thinking controls and default:
llama.cpp
sin firmar La última versión, v0.6.0, no lleva ninguna firma que GitHub haya podido verificar.
Cargando el informe de seguridad
Ollama
✓ firmada La última versión, v0.35.1, lleva una firma que GitHub ha verificado.
Cargando el informe de seguridad