Run frontier AI locally.
Splits one model across several of your devices, behind OpenAI, Claude and Ollama compatible APIs.
- Lenguaje
- Python
- Licencia
- Apache-2.0
- Estrellas
- 48k
- Última versión
- v1.0.71
- Último push
- 2026-10-06
3 herramientas sustituyen a Ollama
Run frontier AI locally.
Splits one model across several of your devices, behind OpenAI, Claude and Ollama compatible APIs.
Distribute and run LLMs with a single file.
Packs llama.cpp and a model into one executable that runs on most systems, with a local API and web chat.
Universal LLM Deployment Engine with ML Compilation
Compiles models for GPUs, phones and browsers, and serves them behind an OpenAI-compatible API.
Cargando el README
✓ firmada La última versión, v0.35.1, lleva una firma que GitHub ha verificado.
Cargando el informe de seguridad