Qwen-Audio-3.1-TTS-Plus
Top-3 voice quality at a quarter of the leader’s price.
Qwen-Audio-3.1-TTS-Plus is Alibaba’s text-to-speech model, ranking third in the Artificial Analysis TTS arena at about $19 per million characters.
- Provider
- Alibaba (Qwen)
- Type
- Text-to-speech
- Released
- Sep 24, 2026
- Price
- ≈$19.31 per 1M characters
- Input
- text
- Output
- audio
- Open weights
- No
Best for
- Multimodal
- High volume / low cost
Strengths
- Top-3 speech arena Elo
- Low price per character
More from Alibaba (Qwen)
- Qwen3.8 Max — Alibaba (Qwen). Alibaba’s 2.4T-parameter proprietary flagship with video understanding.
- Qwen3.8 2.4T A95B — Alibaba (Qwen). The open-weight variant of Qwen3.8 Max — the largest open model available.
- Qwen3.8 Flash — Alibaba (Qwen). Open multimodal reasoning model for coding, desktop interaction and long video.
- Qwen3.6 35B A3B — Alibaba (Qwen). Tiny-active-parameter open MoE — a local-inference favorite.
- Qwen3-Embedding-8B — Alibaba (Qwen). Leading open-weight multilingual embedding model.
- Qwen-Image-3.0-Pro — Alibaba (Qwen). Alibaba’s flagship image model with strong text rendering.
- Qwen-Image-2.1 — Alibaba (Qwen). Highest-ranked open-weight image model in the arena.
- Wan 3.0 — Alibaba (Qwen). #1 in the Artificial Analysis text-to-video arena.