← volver

categoría

Model serving

Serve trained machine learning models behind an API, with batching, scaling and versioning.

Comparativa

Model serving
HerramientaEstrellasLicenciaCondicionesAutoalojableLenguajeÚltima versiónÚltimo push
Triton Inference Server11kBSD-3-ClauseCódigo abiertoSíPythonv2.73.02026-10-06
BentoML8.9kApache-2.0Código abiertoSíPythonv1.4.392026-10-05
KServe6.1kApache-2.0Código abiertoSíGov0.21.02026-10-06
MLServer900Apache-2.0Código abiertoSíPython1.7.12026-10-05

4 herramientas

BentoML

The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!

Lenguaje
Python
Licencia
Apache-2.0
Estrellas
8.9k
Última versión
v1.4.39
Último push
2026-10-05
KServe

Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes

Lenguaje
Go
Licencia
Apache-2.0
Estrellas
6.1k
Última versión
v0.21.0
Último push
2026-10-06
MLServer

An inference server for your machine learning models, including support for multiple frameworks, multi-model serving and more

Lenguaje
Python
Licencia
Apache-2.0
Estrellas
900
Última versión
1.7.1
Último push
2026-10-05

Lo que sustituyen estas herramientas