The Triton Inference Server provides an optimized cloud and edge inferencing solution.
- Sprache
- Python
- Lizenz
- BSD-3-Clause
- Sterne
- 11k
- Neuestes Release
- v2.73.0
- Letzter Push
- 2026-10-06
Serve trained machine learning models behind an API, with batching, scaling and versioning.
Änderungen in dieser Kategorie (RSS)
| Tool | Sterne | Lizenz | Bedingungen | Selbst hostbar | Sprache | Neuestes Release | Letzter Push |
|---|---|---|---|---|---|---|---|
| Triton Inference Server | 11k | BSD-3-Clause | Open Source | Ja | Python | v2.73.0 | 2026-10-06 |
| BentoML | 8.9k | Apache-2.0 | Open Source | Ja | Python | v1.4.39 | 2026-10-05 |
| KServe | 6.1k | Apache-2.0 | Open Source | Ja | Go | v0.21.0 | 2026-10-06 |
| MLServer | 900 | Apache-2.0 | Open Source | Ja | Python | 1.7.1 | 2026-10-05 |
4 Tools
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
An inference server for your machine learning models, including support for multiple frameworks, multi-model serving and more