Released on Jul 21, 2026
Alibaba Token Plan announced this model on 2026-07-21.
- Hosting route
- First-party API
- Affected scope
- Not stated in the notice
- Announced
- Jul 21, 2026
- First seen by ModelClock
- Oct 6, 2026
What the provider published
alibabacloud.com ↗July 21, 2026
11,074
0
Our text-to-speech model, now across 16 languages.
Qwen-Audio-3.0-TTS is our latest text-to-speech model release. It ships as two variants from the same lineage:
Flash : tuned for real-time interaction, with a first-packet latency at 300ms-level.
Plus : tuned for high-quality generation, where naturalness and timbre fidelity matter more than speed.
API: Model Studio
Release Blog: Blog
This release focuses on four things developers actually run into in production: broader language coverage, natural-language style control, fine-grained tag control, and robustness when the reference audio isn’t clean.
Qwen-Audio-3.0-TTS-Plus currently ranks #1 on Artificial Analysis, the independent third-party TTS leaderboard.
Read from the provider's text; the quote is the provider's exact lines.