The Triton Inference Server provides an optimized cloud and edge inferencing solution.
- Langage
- Python
- Licence
- BSD-3-Clause
- Étoiles
- 11k
- Dernière release
- v2.73.0
- Dernier push
- 2026-10-06
Serve trained machine learning models behind an API, with batching, scaling and versioning.
Changements dans cette catégorie (RSS)
| Outil | Étoiles | Licence | Conditions | Auto-hébergeable | Langage | Dernière release | Dernier push |
|---|---|---|---|---|---|---|---|
| Triton Inference Server | 11k | BSD-3-Clause | Open source | Oui | Python | v2.73.0 | 2026-10-06 |
| BentoML | 8.9k | Apache-2.0 | Open source | Oui | Python | v1.4.39 | 2026-10-05 |
| KServe | 6.1k | Apache-2.0 | Open source | Oui | Go | v0.21.0 | 2026-10-06 |
| MLServer | 900 | Apache-2.0 | Open source | Oui | Python | 1.7.1 | 2026-10-05 |
4 outils
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
An inference server for your machine learning models, including support for multiple frameworks, multi-model serving and more