llm leaderboard

Best AI for text to speech

Blind pairwise listening votes. Models still marked preliminary upstream are left out until their rating settles.

Aldena runs these models inside your team rooms. See what each one costs.

34 of 34 ranked models
Ranked models
rankmodelvendorcompositebenchmarkselo
1
Luna TTSvuilabs/luna-tts
Vui Labs66.71 of 1 benchmarks1574
2
Luck Dolphinluck-dolphin/luck-dolphin
Luck Dolphin65.71 of 1 benchmarks1561
3
CastleFlow v1.0async/async-1
Async64.61 of 1 benchmarks1560
4
Inworld TTS (max)inworld/inworld:max
Inworld63.61 of 1 benchmarks1558
5
Luck Dolphin Turboluck-dolphin/luck-dolphin-turbo
Luck Dolphin62.61 of 1 benchmarks1550
6
Papla P1papla/papla-p1
Papla61.61 of 1 benchmarks1549
7
Inworld TTSinworld/inworld
Inworld60.61 of 1 benchmarks1543
8
Speech 2.8 HDminimax/speech-2.8-hd
MiniMax59.61 of 1 benchmarks1536
9
Lightning v3.1 Prolightning/lightning-v3.1-pro
Lightning58.61 of 1 benchmarks1530
10
Hume Octavehume/hume-octave
Hume57.61 of 1 benchmarks1527
11
Gradium TTSgradium/gradium
Gradium56.61 of 1 benchmarks1525
12
Speech 2.8 Turbominimax/speech-2.8-turbo
MiniMax55.61 of 1 benchmarks1523
13
OpenAudio S2lanternfish/lanternfish-2
Lanternfish54.51 of 1 benchmarks1522
14
MiniMax Speech 2.6 Turbominimax/minimax-speech-2.6-turbo
MiniMax53.51 of 1 benchmarks1520
15
Cartesia Sonic 2cartesia/cartesia-sonic-2
Cartesia52.51 of 1 benchmarks1517
16
MiniMax Speech 2.6 HDminimax/minimax-speech-2.6-hd
MiniMax51.01 of 1 benchmarks1514
17
Typecast SSFM 3.0typecast/typecast
Typecast51.01 of 1 benchmarks1514
18
MiniMax Speech 02 HDminimax/minimax-speech-02-hd
MiniMax49.51 of 1 benchmarks1513
19
OpenAudio S1lanternfish/lanternfish-1
Lanternfish48.51 of 1 benchmarks1512
20
Eleven Turbo v2.5elevenlabs/eleven-turbo-v2.5
ElevenLabs46.51 of 1 benchmarks1511
21
Hithink Speech 2.6hithink/hithink-speech-2.6
Hithink46.51 of 1 benchmarks1511
22
Inworld TTS 1.5 MAXinworld/inworld-max-1.5
Inworld46.51 of 1 benchmarks1511
23
Eleven Flash v2.5elevenlabs/eleven-flash-v2.5
ElevenLabs43.41 of 1 benchmarks1508
24
Eleven Multilingual v2elevenlabs/eleven-multilingual-v2
ElevenLabs43.41 of 1 benchmarks1508
25
Eleven v3elevenlabs/eleven-v3
ElevenLabs43.41 of 1 benchmarks1508
26
MiniMax Speech 02 Turbominimax/minimax-speech-02-turbo
MiniMax41.41 of 1 benchmarks1504
27
Voice.ai Text to Speech V1voiceai/voiceai-tts-v1
Voice.ai40.41 of 1 benchmarks1498
28
Chatterboxresemble-ai/chatterbox
Resemble AI39.41 of 1 benchmarks1480
29
Kokoro v1.0hexgrad/kokoro-v1
Hexgrad38.41 of 1 benchmarks1478
30
Magpie Research Previewmagpie/magpie-rp
Magpie37.41 of 1 benchmarks1474
31
NeuTTS Maxneuphonic/neuphonic
Neuphonic36.41 of 1 benchmarks1440
32
Maya 1maya-research/maya1
Maya Research35.41 of 1 benchmarks1411
33
Magpie Multilingualmagpie/magpie
Magpie34.31 of 1 benchmarks1406
34
Wordcab TTSwordcab/wordcab
Wordcab33.31 of 1 benchmarks1361
How this ranks

Every benchmark value becomes a percentile among the models that have it, so accuracy scores, Elo ratings and word error rates compare without hand-tuned scaling. Metrics where lower is better are inverted first. Raw values are never summed or averaged across benchmarks. A model's mean percentile is then shrunk toward the mean of the models that were broadly benchmarked, so a model tested twice cannot outrank a broadly tested one on two lucky results. Turning a data source off runs that same ranking code again in your browser over the sources you left on.

A model scored on fewer than 1 of the 1 ranked benchmarks in this category still ranks here, on the benchmarks it does have, and its row carries a partial coverage mark. On an equal score it sits under the model that earned the same number across more of the board.

Data sources

Turn a source off to drop every benchmark it feeds and rank the board again from what is left, in your browser. Turn them all off and the table has nothing to rank. Your choice follows you across the leaderboard pages.