本文へスキップ
>
awesome-alternatives
カタログを検索
→
一覧
概要
ツールを追加
言語
English
Français
Español
Deutsch
Português (Brasil)
日本語
Bahasa Indonesia
English
Français
Español
Deutsch
Português (Brasil)
日本語
Bahasa Indonesia
ホーム
Local model runtimes
LMDeploy
代替対象:
ChatGPT
LMDeploy
README
リリース
5
セキュリティ
置き換え対象
1
README を読み込み中
v0.18.0
2026-09-28
[Feat]: Support output input logprobs
v0.17.0
2026-09-01
Integrate DeepEPv2
v0.16.0
2026-08-19
Support Interns2 mobius
v0.15.0
2026-07-31
Support long-context and MTP prefix-cache hits
v0.14.0
2026-06-24
FP8 kv cache quantization
✓ 署名済み
最新リリース v0.18.0 には、GitHub が検証した署名があります。
セキュリティレポートを読み込み中
ChatGPT
Serves open-weight models behind an OpenAI-compatible API, with quantisation and batching.
部分的
Local model runtimesのその他のツール
Ollama
Go
→
llama.cpp
C++
→
vLLM
Python
→
LocalAI
Go
→
exo
Python
→
SGLang
Python
→
llamafile
C++
→
MLC LLM
Python
→
KTransformers
Python
→
TensorRT-LLM
Python
→
Xinference
Python
→
mistral.rs
Rust
→
llama-swap
Go
→
Lemonade
C++
→
GPUStack
Python
→
RamaLama
Python
→
TabbyAPI
Python
→