← retour

catégorie

Model serving

Serve trained machine learning models behind an API, with batching, scaling and versioning.

Côte à côte

Model serving
OutilÉtoilesLicenceConditionsAuto-hébergeableLangageDernière releaseDernier push
Triton Inference Server11kBSD-3-ClauseOpen sourceOuiPythonv2.73.02026-10-06
BentoML8.9kApache-2.0Open sourceOuiPythonv1.4.392026-10-05
KServe6.1kApache-2.0Open sourceOuiGov0.21.02026-10-06
MLServer900Apache-2.0Open sourceOuiPython1.7.12026-10-05

4 outils

BentoML

The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!

Langage
Python
Licence
Apache-2.0
Étoiles
8.9k
Dernière release
v1.4.39
Dernier push
2026-10-05
KServe

Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes

Langage
Go
Licence
Apache-2.0
Étoiles
6.1k
Dernière release
v0.21.0
Dernier push
2026-10-06
MLServer

An inference server for your machine learning models, including support for multiple frameworks, multi-model serving and more

Langage
Python
Licence
Apache-2.0
Étoiles
900
Dernière release
1.7.1
Dernier push
2026-10-05

Ce que ces outils remplacent