compare
Chatterbox vs F5-TTS
The same facts for both, read from GitHub every night, and the relation a person reviewed.
| Fact | Chatterbox | F5-TTS |
|---|---|---|
| Language | Python | Python |
| Licence | MIT | MIT |
| Stars | 27k | 15k |
| Latest | v0.1.2 | 1.1.22 |
| Last push | 2026-07-21 | 2026-09-21 |
| Release cadence | too few releases | about 20 days between releases |
| Active contributors | 1 commit author on the default branch in the last 90 days | 3 commit authors on the default branch in the last 90 days |
| Flags | none | none |
How they relate
Both replace ElevenLabs. Alternatives to ElevenLabs →
PartialChatterboxAn open TTS model with zero-shot voice cloning and emotion control.
PartialF5-TTSVoice cloning from a short reference clip, with a Gradio app.
Chatterbox
- v0.1.22025-06-13Example for use in mac M*
F5-TTS
- 1.1.222026-07-23fix: allocate fix_duration across chunks to avoid N× duration blow-up
- 1.1.212026-07-05update gradio>=6.15.0 in pyproject.toml
- 1.1.202026-04-20fix: refactor cache handling in DiT, MMDiT, and UNetT classes (lazyinit to avoid EMA deepcopy failure while training)
- 1.1.192026-04-16reuse resamplers and cache vocos MelSpectrogram instances
- 1.1.182026-03-24Add Arabic model details to SHARED.md
Chatterbox
✓ signed The latest release, v0.1.2, carries a signature GitHub verified.
Loading the security report
F5-TTS
✓ signed The latest release, 1.1.22, carries a signature GitHub verified.
Loading the security report