OpenSpeech — Open-source TTS, side-by-side

OpenSpeech

1 min read Original article ↗

Start here

Picks updated by community votes

Kokoro-82M

The current default recommendation.

Excellent quality. Tiny model. Runs almost anywhere.

Orpheus TTS

Most natural-sounding conversations.

Llama-3 based. Empathetic. Built for interactive apps.

Chatterbox

Best voice cloning, MIT-licensed.

ElevenLabs-tier quality from a few seconds of reference audio.

The three scripts (every voice reads these)

Neutral

The quick brown fox jumps over the lazy dog near the riverbank.

Baseline naturalness — no emotion, no tricky words

Emotional

I can't believe you actually did it. This is incredible!

Prosody and excitement — tests expressiveness

Numbers & Dates

On March 14th, 2025, the team raised $4.2 million at a 38% margin.

Number, date, and symbol pronunciation — the common failure mode

Browse by need

Not sure what you need yet?

Each shelf below is hand-curated. Scroll horizontally to see all picks. Add anything to the comparison bar to hear them side-by-side.

Best for voice cloning

Clone a voice from a short reference clip.

Best for real-time

Stream tokens fast enough for voice agents.

Best for multilingual

Five or more languages in one model.

Best for edge deployment

Under 200M parameters — embed anywhere.

Best for emotion + flavor

Long-form, dialogue, and expressive models.

Want the full directory?

All 26 models with full filters and sort. 46 voices total. For when you're ready to do your own research.

Browse all 26 models