alternativas a

ElevenLabs

Hosted text-to-speech and voice cloning service, with an API. Un producto cerrado de ElevenLabs.

Comparativa

9 alternativas a ElevenLabs
HerramientaSustituciónEstrellasLicenciaCondicionesAutoalojableLenguajeÚltima versiónÚltimo push
GPT-SoVITSParcial62kMITCódigo abiertoSíPython20250606v2pro2026-10-06
Fish SpeechParcial33kOtherCódigo disponibleSíPythonv1.5.12026-10-05
ChatterboxParcial27kMITCódigo abiertoSíPythonv0.1.22026-07-21
IndexTTSParcial24kOtherCódigo disponibleSíPythonv2.5.02026-09-29
CosyVoiceParcial24kApache-2.0Código abiertoSíPythonv2.02026-05-25
ebook2audiobookParcial20kApache-2.0Código abiertoSíPythonv26.10.12026-10-06
F5-TTSParcial15kMITCódigo abiertoSíPython1.1.222026-09-21
PiperParcial5.8kGPL-3.0Código abiertoSíC++v1.8.02026-09-28
Kokoro-FastAPIParcial5.5kApache-2.0Código abiertoSíPythonv0.9.02026-10-05

Todas las alternativas

Lenguaje
Licencia
Condiciones
Despliegue

9 alternativas

GPT-SoVITSParcial

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

Voice cloning and fine-tuning from a minute of audio, with a web UI.

Lenguaje
Python
Licencia
MIT
Estrellas
62k
Última versión
20250606v2pro
Último push
2026-10-06
Fish SpeechParcial

SOTA Open Source TTS

Multilingual TTS with voice cloning, under a non-commercial licence.

Lenguaje
Python
Licencia
Other
Estrellas
33k
Última versión
v1.5.1
Último push
2026-10-05
ChatterboxParcial

SoTA open-source TTS

An open TTS model with zero-shot voice cloning and emotion control.

Lenguaje
Python
Licencia
MIT
Estrellas
27k
Última versión
v0.1.2
Último push
2026-07-21
IndexTTSParcial

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Zero-shot TTS with voice cloning and control over duration and emotion.

Lenguaje
Python
Licencia
Other
Estrellas
24k
Última versión
v2.5.0
Último push
2026-09-29
CosyVoiceParcial

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Multilingual TTS with zero-shot voice cloning and streaming output.

Lenguaje
Python
Licencia
Apache-2.0
Estrellas
24k
Última versión
v2.0
Último push
2026-05-25
ebook2audiobookParcial

Generate audiobooks from e-books, voice cloning & 1158+ languages!

Converts ebooks into audiobooks with local TTS models, including cloned voices.

Lenguaje
Python
Licencia
Apache-2.0
Estrellas
20k
Última versión
v26.10.1
Último push
2026-10-06
F5-TTSParcial

Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"

Voice cloning from a short reference clip, with a Gradio app.

Lenguaje
Python
Licencia
MIT
Estrellas
15k
Última versión
1.1.22
Último push
2026-09-21
PiperParcial

Fast and local neural text-to-speech engine

Fast offline voices that run on a Raspberry Pi, without voice cloning.

Lenguaje
C++
Licencia
GPL-3.0
Estrellas
5.8k
Última versión
v1.8.0
Último push
2026-09-28
Kokoro-FastAPIParcial

Dockerized OpenAI-compatible wrapper for Kokoro-82M text-to-speech w/multiplatform CPU, AMD, NVIDIA GPU PyTorch; multi-speaker, clone-tuning, caption timestamps, SSML, optional readalong web UI

An OpenAI-compatible speech API around the Kokoro model, on CPU or GPU, without voice cloning.

Lenguaje
Python
Licencia
Apache-2.0
Estrellas
5.5k
Última versión
v0.9.0
Último push
2026-10-05

Cara a cara

Sobre ElevenLabs

Empresa
ElevenLabs

Web de ElevenLabs ↗