Aller au contenu

Alibaba Qwen-Audio-3.0-TTS-Plus Alibaba (Qwen) · text-to-speech (TTS) API

Listed

Speech Arena Elo 1,259 on Artificial Analysis, latency not measured; ranked 40th of 112 in the TTS for Voice Agents 2026 index.

Benchmark summary: 11.1/100, ranked 40th of 112 in the TTS for Voice Agents 2026 index (Listed). Strongest criterion: voice quality, 9.7 of 20 points.

Alibaba Qwen-Audio-3.0-TTS-Plus is a proprietary text-to-speech (TTS) API from Alibaba (Qwen), measured by one independent benchmark, Artificial Analysis.

Overall score

11.1 /100

40 th of 112

Score breakdown

  • Real-time speed 0 /60
  • Voice quality 9.7 /20
  • Accuracy 0 /15
  • Price 1.4 /5

Data checked on September 30, 2026

Updated after each benchmark capture

Every point comes from a public benchmark: sources · methodology

Gradium TTS Beta vs Alibaba Qwen-Audio-3.0-TTS-Plus Compare with VUI Labs Luna TTS

Alibaba Qwen-Audio-3.0-TTS-Plus ranks 40th of 112 in the 2026 text-to-speech (TTS) API benchmark, with a score of 11.1/100.

Real-time latency

Neither Coval nor Speko measures Alibaba Qwen-Audio-3.0-TTS-Plus's latency, so it scores 0 of 60 on real-time speed and does not appear in the real-time ranking.

Accuracy (word error rate)

No accuracy measurement: Coval does not measure its word error rate and Speko does not rate its pronunciation robustness, so accuracy counts 0 of 15.

Voice quality

In the Artificial Analysis Speech Arena (capture of October 1, 2026), Alibaba Qwen-Audio-3.0-TTS-Plus has an Elo of 1,259 (±16, 95% confidence) from 1,747 blind comparisons, 4th highest of the 92 models Artificial Analysis rates.

Price

Artificial Analysis lists Alibaba Qwen-Audio-3.0-TTS-Plus at $19.3 per 1M characters (34th lowest of the 78 models Artificial Analysis lists a price for).

Scoring note

Artificial Analysis changed its Speech Arena Elo between the captures of September 30, 2026 and October 1, 2026 without a new measurement date (1,258, then 1,259): the score uses the average, 1,258.5. The value shown is the latest.

Languages tested

Tested in Spanish, German, French, Portuguese, Japanese, Chinese, Arabic and Hindi by Artificial Analysis or Speko. A language missing here was not tested, which does not mean it is unsupported.

Benchmark measurements

Artificial Analysis Speech Arena

Speech Arena Elo (blind listener preference)
1,259
Elo 95% interval (±)
16
Arena rank
4
Arena rank range
2-5
Arena appearances
1,747
Voices tested
8
Price per 1M characters
19.3 USD
Released
Sept 2026
Artificial Analysis model
Qwen-Audio-3.0-TTS-Plus
Artificial Analysis leaderboard
artificialanalysis.ai

Not collected

Similar TTS models