Best AI for text to speech
Blind pairwise listening votes. Models still marked preliminary upstream are left out until their rating settles.
Aldena runs these models inside your team rooms. See what each one costs.
| rank | model | vendor | composite | benchmarks | elo |
|---|---|---|---|---|---|
| 1 | Luna TTSvuilabs/luna-tts | Vui Labs | 66.7 | 1 of 1 benchmarks | 1574 |
| 2 | Luck Dolphinluck-dolphin/luck-dolphin | Luck Dolphin | 65.7 | 1 of 1 benchmarks | 1561 |
| 3 | CastleFlow v1.0async/async-1 | Async | 64.6 | 1 of 1 benchmarks | 1560 |
| 4 | Inworld TTS (max)inworld/inworld:max | Inworld | 63.6 | 1 of 1 benchmarks | 1558 |
| 5 | Luck Dolphin Turboluck-dolphin/luck-dolphin-turbo | Luck Dolphin | 62.6 | 1 of 1 benchmarks | 1550 |
| 6 | Papla P1papla/papla-p1 | Papla | 61.6 | 1 of 1 benchmarks | 1549 |
| 7 | Inworld TTSinworld/inworld | Inworld | 60.6 | 1 of 1 benchmarks | 1543 |
| 8 | Speech 2.8 HDminimax/speech-2.8-hd | MiniMax | 59.6 | 1 of 1 benchmarks | 1536 |
| 9 | Lightning v3.1 Prolightning/lightning-v3.1-pro | Lightning | 58.6 | 1 of 1 benchmarks | 1530 |
| 10 | Hume Octavehume/hume-octave | Hume | 57.6 | 1 of 1 benchmarks | 1527 |
| 11 | Gradium TTSgradium/gradium | Gradium | 56.6 | 1 of 1 benchmarks | 1525 |
| 12 | Speech 2.8 Turbominimax/speech-2.8-turbo | MiniMax | 55.6 | 1 of 1 benchmarks | 1523 |
| 13 | OpenAudio S2lanternfish/lanternfish-2 | Lanternfish | 54.5 | 1 of 1 benchmarks | 1522 |
| 14 | MiniMax Speech 2.6 Turbominimax/minimax-speech-2.6-turbo | MiniMax | 53.5 | 1 of 1 benchmarks | 1520 |
| 15 | Cartesia Sonic 2cartesia/cartesia-sonic-2 | Cartesia | 52.5 | 1 of 1 benchmarks | 1517 |
| 16 | MiniMax Speech 2.6 HDminimax/minimax-speech-2.6-hd | MiniMax | 51.0 | 1 of 1 benchmarks | 1514 |
| 17 | Typecast SSFM 3.0typecast/typecast | Typecast | 51.0 | 1 of 1 benchmarks | 1514 |
| 18 | MiniMax Speech 02 HDminimax/minimax-speech-02-hd | MiniMax | 49.5 | 1 of 1 benchmarks | 1513 |
| 19 | OpenAudio S1lanternfish/lanternfish-1 | Lanternfish | 48.5 | 1 of 1 benchmarks | 1512 |
| 20 | Eleven Turbo v2.5elevenlabs/eleven-turbo-v2.5 | ElevenLabs | 46.5 | 1 of 1 benchmarks | 1511 |
| 21 | Hithink Speech 2.6hithink/hithink-speech-2.6 | Hithink | 46.5 | 1 of 1 benchmarks | 1511 |
| 22 | Inworld TTS 1.5 MAXinworld/inworld-max-1.5 | Inworld | 46.5 | 1 of 1 benchmarks | 1511 |
| 23 | Eleven Flash v2.5elevenlabs/eleven-flash-v2.5 | ElevenLabs | 43.4 | 1 of 1 benchmarks | 1508 |
| 24 | Eleven Multilingual v2elevenlabs/eleven-multilingual-v2 | ElevenLabs | 43.4 | 1 of 1 benchmarks | 1508 |
| 25 | Eleven v3elevenlabs/eleven-v3 | ElevenLabs | 43.4 | 1 of 1 benchmarks | 1508 |
| 26 | MiniMax Speech 02 Turbominimax/minimax-speech-02-turbo | MiniMax | 41.4 | 1 of 1 benchmarks | 1504 |
| 27 | Voice.ai Text to Speech V1voiceai/voiceai-tts-v1 | Voice.ai | 40.4 | 1 of 1 benchmarks | 1498 |
| 28 | Chatterboxresemble-ai/chatterbox | Resemble AI | 39.4 | 1 of 1 benchmarks | 1480 |
| 29 | Kokoro v1.0hexgrad/kokoro-v1 | Hexgrad | 38.4 | 1 of 1 benchmarks | 1478 |
| 30 | Magpie Research Previewmagpie/magpie-rp | Magpie | 37.4 | 1 of 1 benchmarks | 1474 |
| 31 | NeuTTS Maxneuphonic/neuphonic | Neuphonic | 36.4 | 1 of 1 benchmarks | 1440 |
| 32 | Maya 1maya-research/maya1 | Maya Research | 35.4 | 1 of 1 benchmarks | 1411 |
| 33 | Magpie Multilingualmagpie/magpie | Magpie | 34.3 | 1 of 1 benchmarks | 1406 |
| 34 | Wordcab TTSwordcab/wordcab | Wordcab | 33.3 | 1 of 1 benchmarks | 1361 |
How this ranks
Every benchmark value becomes a percentile among the models that have it, so accuracy scores, Elo ratings and word error rates compare without hand-tuned scaling. Metrics where lower is better are inverted first. Raw values are never summed or averaged across benchmarks. A model's mean percentile is then shrunk toward the mean of the models that were broadly benchmarked, so a model tested twice cannot outrank a broadly tested one on two lucky results. Turning a data source off runs that same ranking code again in your browser over the sources you left on.
A model scored on fewer than 1 of the 1 ranked benchmarks in this category still ranks here, on the benchmarks it does have, and its row carries a partial coverage mark. On an equal score it sits under the model that earned the same number across more of the board.
Data sources
Turn a source off to drop every benchmark it feeds and rank the board again from what is left, in your browser. Turn them all off and the table has nothing to rank. Your choice follows you across the leaderboard pages.
- TTS Arena V2
Elo ratings from TTS Arena V2.