Microsoft MAI-Voice-1
Microsoft MAI-Voice-1 is a proprietary text-to-speech (TTS) API from Microsoft, measured by one independent benchmark, Artificial Analysis.
Speech Arena Elo 1,031 on Artificial Analysis, latency not measured; ranked 97th (tied) of 112 in the TTS for Voice Agents 2026 index.
Benchmark summary: 2.8/100, ranked 97th (tied) of 112 in the TTS for Voice Agents 2026 index (Listed). Strongest criterion: voice quality, 2.6 of 20 points.
Amazon Polly Long-Form is a proprietary text-to-speech (TTS) API from Amazon, measured by one independent benchmark, Artificial Analysis.
Overall score
2.8 /100
Data checked on September 30, 2026
Updated after each benchmark capture
Every point comes from a public benchmark: sources · methodology
Amazon Polly Long-Form ranks 97th of 112 in the 2026 text-to-speech (TTS) API benchmark, with a score of 2.8/100.
Neither Coval nor Speko measures Amazon Polly Long-Form's latency, so it scores 0 of 60 on real-time speed and does not appear in the real-time ranking.
No accuracy measurement: Coval does not measure its word error rate and Speko does not rate its pronunciation robustness, so accuracy counts 0 of 15.
In the Artificial Analysis Speech Arena (capture of October 1, 2026), Amazon Polly Long-Form has an Elo of 1,031 (±13, 95% confidence) from 1,893 blind comparisons, 68th highest of the 92 models Artificial Analysis rates.
Artificial Analysis lists Amazon Polly Long-Form at $100 per 1M characters (69th lowest of the 78 models Artificial Analysis lists a price for).
Artificial Analysis changed its Speech Arena Elo between the captures of September 30, 2026 and October 1, 2026 without a new measurement date (1,032, then 1,031): the score uses the average, 1,031.5. The value shown is the latest.
Microsoft MAI-Voice-1 is a proprietary text-to-speech (TTS) API from Microsoft, measured by one independent benchmark, Artificial Analysis.
Alibaba Qwen3 TTS is a proprietary text-to-speech (TTS) API from Alibaba (Qwen), measured by one independent benchmark, Artificial Analysis.
StyleTTS 2 is an open-weights text-to-speech (TTS) API from StyleTTS, measured by one independent benchmark, Artificial Analysis.