1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Voice cloning and fine-tuning from a minute of audio, with a web UI.
- Bahasa
- Python
- Lisensi
- MIT
- Bintang
- 62k
- Terbaru
- 20250606v2pro
- Push terakhir
- 2026-10-06
alternatif untuk
Hosted text-to-speech and voice cloning service, with an API. Produk tertutup dari ElevenLabs.
| Alat | Kecocokan | Bintang | Lisensi | Ketentuan | Self-hosted | Bahasa | Terbaru | Push terakhir |
|---|---|---|---|---|---|---|---|---|
| GPT-SoVITS | Sebagian | 62k | MIT | Open source | Ya | Python | 20250606v2pro | 2026-10-06 |
| Fish Speech | Sebagian | 33k | Other | Source available | Ya | Python | v1.5.1 | 2026-10-05 |
| Chatterbox | Sebagian | 27k | MIT | Open source | Ya | Python | v0.1.2 | 2026-07-21 |
| IndexTTS | Sebagian | 24k | Other | Source available | Ya | Python | v2.5.0 | 2026-09-29 |
| CosyVoice | Sebagian | 24k | Apache-2.0 | Open source | Ya | Python | v2.0 | 2026-05-25 |
| ebook2audiobook | Sebagian | 20k | Apache-2.0 | Open source | Ya | Python | v26.10.1 | 2026-10-06 |
| F5-TTS | Sebagian | 15k | MIT | Open source | Ya | Python | 1.1.22 | 2026-09-21 |
| Piper | Sebagian | 5.8k | GPL-3.0 | Open source | Ya | C++ | v1.8.0 | 2026-09-28 |
| Kokoro-FastAPI | Sebagian | 5.5k | Apache-2.0 | Open source | Ya | Python | v0.9.0 | 2026-10-05 |
9 alternatif
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Voice cloning and fine-tuning from a minute of audio, with a web UI.
SOTA Open Source TTS
Multilingual TTS with voice cloning, under a non-commercial licence.
SoTA open-source TTS
An open TTS model with zero-shot voice cloning and emotion control.
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
Zero-shot TTS with voice cloning and control over duration and emotion.
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Multilingual TTS with zero-shot voice cloning and streaming output.
Generate audiobooks from e-books, voice cloning & 1158+ languages!
Converts ebooks into audiobooks with local TTS models, including cloned voices.
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
Voice cloning from a short reference clip, with a Gradio app.
Fast and local neural text-to-speech engine
Fast offline voices that run on a Raspberry Pi, without voice cloning.
Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, clone-tuning, caption timestamps, SSML, optional readalong web UI
An OpenAI-compatible speech API around the Kokoro model, on CPU or GPU, without voice cloning.