Last refresh: Sep 6, 2026
texttospeech.com Get the data

The TTS Meta-Leaderboard

The best text-to-speech APIs, ranked

Simba 3.2 by SpeechifyAI is currently the best text-to-speech API, scoring 95/100 across all 3 public TTS benchmarks at $10.00 per 1M characters. We rank 107 models from 46 providers by aggregating artificialanalysis.com, voicearena.com and humannessindex.vapi.ai into one score — a fixed formula, no editorial adjustments, no vendor weighting. The cheapest scored model is Kokoro 82M v1.0 at $0.70 per 1M characters.

Quick answers

Model rankings

Score = Bayesian pairwise rank aggregation across 3 boards — fractional ranks, breadth-aware, chance-proof

RankModel AA — artificialanalysis.comVA — voicearena.comHI — humannessindex.vapi.ai Score $/1M chars
01 Simba 3.2 SpeechifyAI 90.881.297.0 95.0 $10.00/1M
02 Sonic 3.6 Cartesia 100.0100.0 85.0 $49.00/1M
03 Sonic 3.5 Cartesia 80.582.671.0 83.0 $49.00/1M
04 Eleven v3 ElevenLabs 76.660.197.0 83.0 $100.00/1M
05 Gemini 3.1 Flash TTS Google 83.893.4 78.0 $18.30/1M
06 Speech 2.8 HD MiniMax 75.793.0 75.0 $100.00/1M
07 up 6 over the past week Fish Audio S2.1 Pro Fish Audio 69.162.086.0 74.0 $15.00/1M
08 new on the board this week Realtime TTS-2 Inworld 93.4 73.0 $20.80/1M
09 down 1 over the past week Qwen-Audio-3.0-TTS-Plus Alibaba 91.0 72.0 $27.60/1M
10 down 1 over the past week Luna TTS VUI Labs 88.2 71.0 $80.00/1M
11 up 1 over the past week Falcon 2 Murf AI 72.963.8 70.0 $10.00/1M
12 new on the board this week Realtime TTS-2 Flash - Research Preview Inworld 86.9 70.0 $10.40/1M
13 down 2 over the past week Breeze TTS 2 BreezeBlue 85.3 69.0 $34.00/1M
14 down 4 over the past week v3 Conversational ElevenLabs 84.2 68.0 $50.00/1M
15 down 1 over the past week Lightning V3.1 Pro (Jul 2026) Smallest.ai 79.949.3 67.0 $19.50/1M
16 new on the board this week StepAudio 2.5 TTS (Aug 2026) StepFun 83.2 67.0 $85.00/1M
17 Soniox TTS Real-Time v2 Soniox 77.5 64.0 $14.20/1M
18 Speech-02-HD MiniMax 62.890.0 63.0 $100.00/1M
19 Speech 2.8 Turbo MiniMax 71.6 61.0 $60.00/1M
20 new on the board this week Gradium TTS (Aug 2026) Gradium 70.9 60.0 $47.20/1M
Unscored — insufficient board data
Unranked Async Flash v1.5 async 69.1 $10.10/1M
Unranked Step TTS 2 (Mar 2026) StepFun 68.5 $40.00/1M
Unranked Fish Audio S2 Pro Fish Audio Open weights 66.338.0 $15.00/1M
Unranked Speech 2.6 HD MiniMax 67.6 $100.00/1M
Unranked Lightning V3.1 Pro TTS (Jun 2026) Smallest.ai 67.4 $19.50/1M
Unranked Azure HD 2.5 Microsoft 67.0 $22.00/1M
Unranked Speech 2.6 Turbo MiniMax 66.1 $60.00/1M
Unranked SpaceXAI TTS SpaceXAI 65.2 $15.00/1M
Unranked Simba 3.0 SpeechifyAI 65.0 $12.50/1M
Unranked Async Pro v1.0 async 64.8 $20.20/1M
Unranked Speech-02-Turbo MiniMax 56.269.0 $60.00/1M
Unranked TTS-1 HD OpenAI 61.5 $30.00/1M
Unranked Step Audio EditX (Mar 2026) StepFun Open weights 61.1
Unranked Turbo v2.5 ElevenLabs 60.2 $50.00/1M
Unranked Speech 2.6 Hithink 56.753.5
Unranked Multilingual v2 ElevenLabs 60.0 $100.00/1M
Unranked Flash v2.5 ElevenLabs 55.868.0 $50.00/1M
Unranked Chatterbox HD Resemble AI 58.6 $40.00/1M
Unranked TTS-1 OpenAI 58.2 $15.00/1M
Unranked Gemini 2.5 Flash Lite TTS Google 56.7 $9.20/1M
Unranked Studio Google 56.0 $160.00/1M
Unranked OpenAudio S1 Fish Audio 55.8 $15.00/1M
Unranked Sonic 3 Cartesia 53.00.066.0 $49.00/1M
Unranked Voxtral TTS Mistral Open weights 55.4 $16.00/1M
Unranked Maya 2 Global Maya Research 51.262.4
Unranked Journey Google 54.7 $160.00/1M
Unranked Maya 2 Flash Maya Research 54.0
Unranked GPT-Realtime-2 OpenAI 53.0 $191.60/1M
Unranked T2A-01-HD MiniMax 52.7 $50.00/1M
Unranked Magpie-Multilingual 357M (Feb 2026) NVIDIA Open weights 52.5
Unranked Polly Generative Amazon 52.1 $30.00/1M
Unranked Simba 1.6 SpeechifyAI 51.9 $10.00/1M
Unranked Kokoro 82M v1.0 Kokoro Open weights 51.4 $0.70/1M
Unranked MiMo-V2.5-TTS Xiaomi 51.2
Unranked Coda Rime 51.0 $50.00/1M
Unranked Chirp 3: HD Google 50.8 $30.00/1M
Unranked Gemini 2.5 Flash TTS (Dec 2025) Google 50.5 $9.20/1M
Unranked Octave 2 Hume AI 50.5 $87.50/1M
Unranked OpenAudio S1 Mini Fish Audio 49.2 $15.00/1M
Unranked Async Flash v1.0 async 48.8 $10.10/1M
Unranked Polly Long-Form Amazon 47.9 $100.00/1M
Unranked Maya1 Maya Research Open weights 47.7
Unranked Higgs Audio V3 TTS Boson AI Open weights 46.6 $15.00/1M
Unranked Bland Speech v3 Bland AI 45.7 $40.00/1M
Unranked MAI-Voice-1 Microsoft 45.7 $16.00/1M
Unranked Simba 1.0 SpeechifyAI 45.7 $10.00/1M
Unranked Grok TTS xAI 64.893.0 $15.00/1M
Unranked T2A-01-Turbo MiniMax 45.5 $30.00/1M
Unranked Gemini 2.5 Pro (Dec 2025) Google 45.1 $18.30/1M
Unranked Azure Neural Microsoft 44.6 $15.00/1M
Unranked Lightning v3.1 Smallest.ai 42.944.0 $25.00/1M
Unranked Octave TTS Hume AI 44.6 $87.50/1M
Unranked MiMo-V2-TTS Xiaomi 43.5
Unranked Chatterbox Resemble AI Open weights 42.9 $25.00/1M
Unranked Raon SpeechLM Krafton 40.3
Unranked Arcana v3 Rime 39.4 $50.00/1M
Unranked Zonos-v0.1 Zyphra Open weights 38.3 $20.00/1M
Unranked Murf Speech Gen 2 Murf AI 34.1 $100.00/1M
Unranked LMNT LMNT 32.8 $49.00/1M
Unranked VibeVoice 7B Microsoft Open weights 31.5
Unranked VibeVoice 1.5B Microsoft 30.9
Unranked OpenVoice v2 OpenVoice Open weights 28.4 $8.30/1M
Unranked Neuphonic TTS Neuphonic 24.541.0 $20.80/1M
Unranked Orpheus Canopy Labs Open weights 89.0
Unranked Magpie Multilingual NVIDIA 26.5
Unranked Grok TTS (Streaming) xAI 87.0 $15.00/1M
Unranked Qwen3 TTS Flash Alibaba 26.3 $10.00/1M
Unranked Realtime TTS 1.5 Max Inworld 78.0 $26.00/1M
Unranked Kugel 3 KugelAudio 22.1
Unranked Turbo v2 ElevenLabs 76.0
Unranked Qwen3 TTS Alibaba 22.1 $20.00/1M
Unranked XTTS v2 Coqui Open weights 21.2 $40.40/1M
Unranked TTS Gradium 60.637.0
Unranked TTS-2 Inworld 71.0
Unranked WaveNet Google 19.7 $16.00/1M
Unranked Mist V2 Rime 16.6 $31.10/1M
Unranked StyleTTS 2 StyleTTS Open weights 14.9 $2.80/1M
Unranked Dragon HD Omni Microsoft 55.4 $22.00/1M
Unranked Sonic 2 Cartesia 67.0
Unranked Neural2 Google 14.4 $16.00/1M
Unranked Polly Neural Amazon 14.0 $16.00/1M
Unranked Standard Google 12.5 $4.00/1M
Unranked Noiz TTS Noiz 9.6
Unranked gpt-4o-mini-tts OpenAI 33.8
Unranked MetaVoice v1 MetaVoice Open weights 3.3
Unranked Sonic Cartesia 0.0
Unranked Polly Standard Amazon 0.0 $4.00/1M
Unranked Aura Deepgram $15.00
Unranked Aura 2 Deepgram $30.00
Unranked Flux TTS Deepgram $45.00
Unranked WellSaid Studio WellSaid Labs $20.00*
Unranked Play 3.0 Mini Play.ht $35.00*
Unranked PlayDialog Play.ht $35.00*
Unranked PlayDialog Turbo Play.ht $35.00*
Unranked Bland Voice Bland AI $15.56*

Per-board cells show the board's own normalized score; "—" = not listed. The Score is a 0–100 Bayesian pairwise rank aggregate: each board contributes a fractional rank (#1 = 1.0, scaled by board size), breadth pulls single-board models toward the field average, and the Borda aggregate rewards models that are strong across every board — Schulze-tie-broken, with a High/Med/Low confidence label on each model. Carets show score-rank movement over the trailing week. Snapshot: Sep 6, 2026.

Text-to-speech APIs: common questions

What is the best text-to-speech API right now?
Simba 3.2 by SpeechifyAI is the top-ranked text-to-speech model on the texttospeech.com meta-leaderboard, scoring 95/100 with high confidence as of Sep 6, 2026. The score aggregates 3 independent public benchmarks — artificialanalysis.com, voicearena.com and humannessindex.vapi.ai — so it reflects agreement between boards rather than one board's view. It is priced at $10.00 per 1M characters. The ranking is recomputed weekly and the full standings are on the leaderboard.
What is the cheapest text-to-speech API?
Kokoro 82M v1.0 by Kokoro is the cheapest scored text-to-speech model at $0.70 per 1M characters, ranked #53 of 107 with a score of 34/100. The cheapest model inside the top ten is Simba 3.2 by SpeechifyAI at $10.00 per 1M characters, so that is the cheapest option we would call competitive on quality. Prices are $ per 1M characters; per-minute pricing is converted to the same unit so every row compares like with like.
What is the best open-weights text-to-speech model?
Fish Audio S2 Pro by Fish Audio is the highest-ranked open-weights model, #23 overall with 57/100 as of Sep 6, 2026. Open-weights models can be self-hosted instead of called through a vendor API, so their running cost depends on your own hardware rather than a per-character rate. The leaderboard can be filtered to open weights only.
How does texttospeech.com score text-to-speech models?
Every score is a Bayesian pairwise rank aggregation (BPRA) over 3 public boards. Each board contributes a model's rank as a fraction — #1 is 1.0, scaled by how many models the board ranks — a Bayesian coverage term pulls models that appear on few boards toward the field average, and an ecosystem-normalized Borda count turns the result into a 0–100 score: #1 on every board is exactly 100, and missing a board caps the ceiling. Ties break by Schulze path strength. We do not run our own listening tests; we aggregate boards that do, and publish the formula and every source snapshot.
How often is the text-to-speech leaderboard updated?
The source boards are polled weekly and the ranking recomputes whenever any of them changes. The current snapshot was captured Sep 6, 2026. Every snapshot is archived and downloadable as JSON, so any past ranking can be reproduced from the same data and the same published formula.
Is the ranking sponsored or influenced by vendors?
No. There is no vendor sponsorship, referral fee, or paid placement, and no code path that can elevate or demote a named vendor: every model goes through the identical formula, the formula is open source, and each published ranking links the exact source snapshots it was computed from. The source boards are also independent of the vendors they rank — that is one of the criteria a board must meet to be aggregated at all.