MiniMax Speech 2.6 HD
MiniMax Speech 2.6 HD is a proprietary text-to-speech (TTS) API from MiniMax, measured by two independent benchmarks: Artificial Analysis and Speko.
Speech Arena Elo 1,133 on Artificial Analysis, latency not measured; ranked 44th (tied) of 112 in the TTS for Voice Agents 2026 index.
Benchmark summary: 9.7/100, ranked 44th (tied) of 112 in the TTS for Voice Agents 2026 index (Listed). Strongest criterion: price, 2.1 of 5 points.
async Flash v1.5 is a proprietary text-to-speech (TTS) API from Async, measured by one independent benchmark, Artificial Analysis.
Overall score
9.7 /100
Data checked on September 30, 2026
Updated after each benchmark capture
Every point comes from a public benchmark: sources · methodology
async Flash v1.5 ranks 44th of 112 in the 2026 text-to-speech (TTS) API benchmark, with a score of 9.7/100.
Neither Coval nor Speko measures async Flash v1.5's latency, so it scores 0 of 60 on real-time speed and does not appear in the real-time ranking.
No accuracy measurement: Coval does not measure its word error rate and Speko does not rate its pronunciation robustness, so accuracy counts 0 of 15.
In the Artificial Analysis Speech Arena (capture of October 1, 2026), async Flash v1.5 has an Elo of 1,133 (±13, 95% confidence) from 2,202 blind comparisons, 23rd highest of the 92 models Artificial Analysis rates.
Artificial Analysis lists async Flash v1.5 at $10.1 per 1M characters (12th lowest of the 78 models Artificial Analysis lists a price for).
Artificial Analysis changed its Speech Arena Elo between the captures of September 30, 2026 and October 1, 2026 without a new measurement date (1,134, then 1,133): the score uses the average, 1,133.5. The value shown is the latest.
MiniMax Speech 2.6 HD is a proprietary text-to-speech (TTS) API from MiniMax, measured by two independent benchmarks: Artificial Analysis and Speko.
BreezeBlue Breeze TTS 2 is an open-weights text-to-speech (TTS) API from BreezeBlue, measured by one independent benchmark, Artificial Analysis.
ElevenLabs Turbo v2.5 is a proprietary text-to-speech (TTS) API from ElevenLabs, measured by two independent benchmarks: Artificial Analysis and Speko.