Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Runs open-weight models on your own machine behind a local API. Pair it with a chat app for the interface.
OpenAI's AI assistant, used through its web, desktop and mobile apps. OpenAIのクローズドな製品です。
| ツール | 置き換えの度合い | スター | ライセンス | 利用条件 | セルフホスト | 言語 | 最新リリース | 最終プッシュ |
|---|---|---|---|---|---|---|---|---|
| Ollama | 部分的 | 182k | MIT | オープンソース | 可 | Go | v0.35.1 | 2026-10-06 |
| Open WebUI | 部分的 | 154k | Other | ソースアベイラブル | 可 | Python | v0.11.4 | 2026-10-06 |
| llama.cpp | 部分的 | 130k | MIT | オープンソース | 可 | C++ | v0.6.0 | 2026-10-06 |
| vLLM | 部分的 | 93k | Apache-2.0 | オープンソース | 可 | Python | v0.31.0 | 2026-10-06 |
| NextChat | 部分的 | 89k | MIT | オープンソース | 可 | TypeScript | v2.16.1 | 2026-08-11 |
| LobeHub | 部分的 | 83k | Other | ソースアベイラブル | 可 | TypeScript | v2.2.18 | 2026-10-06 |
| AnythingLLM | 部分的 | 67k | MIT | オープンソース | 可 | JavaScript | v1.17.0 | 2026-10-06 |
| Cherry Studio | 部分的 | 52k | AGPL-3.0 | オープンソース | 可 | TypeScript | v2.1.4 | 2026-10-06 |
| LocalAI | 部分的 | 49k | MIT | オープンソース | 可 | Go | v4.11.0 | 2026-10-06 |
| TextGen | 部分的 | 48k | AGPL-3.0 | オープンソース | 可 | Python | v4.9 | 2026-08-17 |
| LibreChat | 部分的 | 45k | MIT | オープンソース | 可 | TypeScript | v0.8.8 | 2026-10-06 |
| Jan | 部分的 | 45k | Other | オープンソース | 可 | Rust | v0.8.4 | 2026-10-06 |
| Fabric | 部分的 | 44k | MIT | オープンソース | 不可 | Go | v1.4.515 | 2026-10-05 |
| Chatbox | 部分的 | 42k | GPL-3.0 | オープンソース | 可 | TypeScript | v1.23.5 | 2026-09-24 |
| Khoj | 部分的 | 38k | AGPL-3.0 | オープンソース | 可 | Python | 2.0.0-beta.28 | 2026-08-02 |
| SGLang | 部分的 | 37k | Apache-2.0 | オープンソース | 可 | Python | v0.5.21 | 2026-10-06 |
| SillyTavern | 部分的 | 34k | AGPL-3.0 | オープンソース | 可 | JavaScript | 1.19.0 | 2026-10-02 |
| TensorRT-LLM | 部分的 | 15k | Other | オープンソース | 可 | Python | v1.2.1 | 2026-10-06 |
| LLM | 部分的 | 13k | Apache-2.0 | オープンソース | 不可 | Python | 0.36 | 2026-09-22 |
| ShellGPT | 部分的 | 12k | MIT | オープンソース | 不可 | Python | 1.5.1 | 2026-07-02 |
| Chat UI | 部分的 | 11k | Apache-2.0 | オープンソース | 可 | TypeScript | v0.10.0 | 2026-10-06 |
| Xinference | 部分的 | 9.6k | Apache-2.0 | オープンソース | 可 | Python | v3.5.0 | 2026-10-06 |
| Page Assist | 部分的 | 8.2k | MIT | オープンソース | 可 | TypeScript | v1.5.86 | 2026-10-04 |
| LMDeploy | 部分的 | 8.1k | Apache-2.0 | オープンソース | 可 | Python | v0.18.0 | 2026-09-28 |
| big-AGI | 部分的 | 7.1k | MIT | オープンソース | 可 | TypeScript | v2.1.0 | 2026-10-05 |
| DeepChat | 部分的 | 6.4k | Apache-2.0 | オープンソース | 可 | TypeScript | v1.1.2 | 2026-10-03 |
| GPUStack | 部分的 | 5.8k | Apache-2.0 | オープンソース | 可 | Python | v2.2.3 | 2026-09-30 |
| 5ire | 部分的 | 5.4k | Other | ソースアベイラブル | 可 | TypeScript | v0.15.4 | 2026-09-14 |
| tgpt | 部分的 | 3.3k | GPL-3.0 | オープンソース | 不可 | Go | v2.15.0 | 2026-10-05 |
| oterm | 部分的 | 2.4k | MIT | オープンソース | 不可 | Python | 0.25.0 | 2026-10-03 |
| Alpaca | 部分的 | 1.6k | GPL-3.0 | オープンソース | 可 | Python | 9.2.5 | 2026-10-02 |
| Hollama | 部分的 | 1.2k | MIT | オープンソース | 可 | TypeScript | 0.36.0 | 2026-10-05 |
代替32件
Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Runs open-weight models on your own machine behind a local API. Pair it with a chat app for the interface.
User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
The chat app, self-hosted, in front of Ollama or any OpenAI-compatible API. The model comes from elsewhere.
LLM inference in C/C++
Runs open-weight models on CPU or GPU, with a built-in server and a minimal web chat.
A high-throughput and memory-efficient inference and serving engine for LLMs
Serves open-weight models behind an OpenAI-compatible API, tuned for GPU throughput.
✨ Zero-config AI chat assistant. No API key needed — sign up and instantly chat with GPT-5, Claude 4, Gemini 2.5, DeepSeek & 100+ top models. Pay-as-you-go saves you more. Available on Web, iOS, macOS, Android, Linux, Windows.
A chat app you deploy yourself in front of OpenAI, Anthropic, Google or a local model. The model comes from elsewhere.
🤯 LobeHub is your Chief Agent Operator, organizing your agents into 7×24 operations by hiring, scheduling, and reporting on your entire AI team.
Chat app with plugins and agents, self-hosted or on the desktop, for most model providers.
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
Desktop or self-hosted app built around chatting with your own documents, with any model provider.
AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs
Desktop app for many hosted providers and local models, with assistants, knowledge bases and MCP. The model comes from elsewhere.
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
OpenAI-compatible API for local text, image and audio models, on CPU or GPU.
Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
Desktop and web app that runs open models locally through llama.cpp and other backends, or calls a hosted API.
Enhanced ChatGPT Clone: Features Agents, MCP, Skills, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching, message search, Code Interpreter, langchain, DALL-E-3, OpenAPI Actions, Functions, Secure Multi-User Auth, Presets, open-source for self-hosting. Active
Self-hosted chat app for Anthropic, OpenAI, Google and local models, with agents and file uploads.
Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.
Desktop app that downloads open models and runs them offline, or calls a hosted API.
Fabric is an open-source framework for augmenting humans using AI. It provides a modular system for solving specific problems using a crowdsourced set of AI prompts that can be used anywhere.
Runs a library of reusable prompts over text piped in from the terminal.
Powerful AI Client
Desktop app in front of OpenAI, Anthropic, Google or a local model through Ollama. The model comes from elsewhere.
Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.
A self-hosted assistant that answers from your documents and the web, in front of a local or hosted model.
SGLang is a high-performance serving framework for large language models and multimodal models.
Serves open-weight models behind an OpenAI-compatible API, tuned for GPU throughput.
LLM Frontend for Power Users.
A self-hosted chat front end aimed at character roleplay, in front of hosted APIs or local models.
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
Serves open-weight models on NVIDIA GPUs behind an OpenAI-compatible API.
Access large language models from the command-line
Prompts hosted and local models from the terminal and logs every exchange to SQLite.
A command-line productivity tool powered by AI large language models like GPT-5, will help you accomplish your tasks faster and more efficiently.
A terminal assistant that answers questions and suggests shell commands.
The open source codebase powering HuggingChat
The chat app behind HuggingChat, self-hosted in front of any OpenAI-compatible API.
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
Serves open-weight language, embedding and speech models behind an OpenAI-compatible API.
Use your locally running AI models to assist you in your web browsing
A browser sidebar and web UI for chatting with local models through Ollama or an OpenAI-compatible API.
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
Serves open-weight models behind an OpenAI-compatible API, with quantisation and batching.
AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. Includes AI personas, AGI functions, world-class Beam multi-model chats, text-to-image, voice, response streaming, code highlighting and execution, PDF import, presets for developers, much more. Deploy on-prem or in the cloud.
Web chat in front of hosted or local models, with answers from several models side by side.
🐬DeepChat - A smart assistant that connects powerful AI to your personal world
Desktop chat app for hosted and local models, with MCP tool calling.
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
Manages a cluster of GPUs and serves open-weight models across it behind an OpenAI-compatible API.
5ire is a cross-platform desktop AI assistant, MCP client. It compatible with major service providers, supports local knowledge base and tools via model context protocol servers .
Desktop chat app for hosted and local models, with MCP tools and a local knowledge base.
AI Chatbots in terminal for free
Terminal chat with several providers, some usable without an API key.
the terminal client for LLMs
A terminal chat client for Ollama models.
🦙 Local and online AI hub
A GNOME desktop app for chatting with local models through Ollama.
A minimal LLM chat app that runs entirely in your browser
A minimal web chat in front of Ollama or an OpenAI-compatible server.