The Triton Inference Server provides an optimized cloud and edge inferencing solution.
Serves TensorRT, PyTorch, ONNX and other models over HTTP and gRPC, with dynamic batching.
- Bahasa
- Python
- Lisensi
- BSD-3-Clause
- Bintang
- 11k
- Terbaru
- v2.73.0
- Push terakhir
- 2026-10-06