1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Voice cloning and fine-tuning from a minute of audio, with a web UI.
- Linguagem
- Python
- Licença
- MIT
- Estrelas
- 62k
- Última release
- 20250606v2pro
- Último push
- 2026-10-06
Hosted text-to-speech and voice cloning service, with an API. Um produto fechado de ElevenLabs.
| Ferramenta | Substituição | Estrelas | Licença | Termos | Auto-hospedado | Linguagem | Última release | Último push |
|---|---|---|---|---|---|---|---|---|
| GPT-SoVITS | Parcial | 62k | MIT | Código aberto | Sim | Python | 20250606v2pro | 2026-10-06 |
| Fish Speech | Parcial | 33k | Other | Código disponível | Sim | Python | v1.5.1 | 2026-10-05 |
| Chatterbox | Parcial | 27k | MIT | Código aberto | Sim | Python | v0.1.2 | 2026-07-21 |
| IndexTTS | Parcial | 24k | Other | Código disponível | Sim | Python | v2.5.0 | 2026-09-29 |
| CosyVoice | Parcial | 24k | Apache-2.0 | Código aberto | Sim | Python | v2.0 | 2026-05-25 |
| ebook2audiobook | Parcial | 20k | Apache-2.0 | Código aberto | Sim | Python | v26.10.1 | 2026-10-06 |
| F5-TTS | Parcial | 15k | MIT | Código aberto | Sim | Python | 1.1.22 | 2026-09-21 |
| Piper | Parcial | 5.8k | GPL-3.0 | Código aberto | Sim | C++ | v1.8.0 | 2026-09-28 |
| Kokoro-FastAPI | Parcial | 5.5k | Apache-2.0 | Código aberto | Sim | Python | v0.9.0 | 2026-10-05 |
9 alternativas
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Voice cloning and fine-tuning from a minute of audio, with a web UI.
SOTA Open Source TTS
Multilingual TTS with voice cloning, under a non-commercial licence.
SoTA open-source TTS
An open TTS model with zero-shot voice cloning and emotion control.
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
Zero-shot TTS with voice cloning and control over duration and emotion.
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Multilingual TTS with zero-shot voice cloning and streaming output.
Generate audiobooks from e-books, voice cloning & 1158+ languages!
Converts ebooks into audiobooks with local TTS models, including cloned voices.
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
Voice cloning from a short reference clip, with a Gradio app.
Fast and local neural text-to-speech engine
Fast offline voices that run on a Raspberry Pi, without voice cloning.
Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, clone-tuning, caption timestamps, SSML, optional readalong web UI
An OpenAI-compatible speech API around the Kokoro model, on CPU or GPU, without voice cloning.