The Triton Inference Server provides an optimized cloud and edge inferencing solution.
- Lenguaje
- Python
- Licencia
- BSD-3-Clause
- Estrellas
- 11k
- Última versión
- v2.73.0
- Último push
- 2026-10-06
Serve trained machine learning models behind an API, with batching, scaling and versioning.
Cambios en esta categoría (RSS)
| Herramienta | Estrellas | Licencia | Condiciones | Autoalojable | Lenguaje | Última versión | Último push |
|---|---|---|---|---|---|---|---|
| Triton Inference Server | 11k | BSD-3-Clause | Código abierto | Sí | Python | v2.73.0 | 2026-10-06 |
| BentoML | 8.9k | Apache-2.0 | Código abierto | Sí | Python | v1.4.39 | 2026-10-05 |
| KServe | 6.1k | Apache-2.0 | Código abierto | Sí | Go | v0.21.0 | 2026-10-06 |
| MLServer | 900 | Apache-2.0 | Código abierto | Sí | Python | 1.7.1 | 2026-10-05 |
4 herramientas
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
An inference server for your machine learning models, including support for multiple frameworks, multi-model serving and more