Introducing the Qwen-Audio-3.0-TTS.
Our latest text-to-speech model, in two flavors:
• Flash: real-time interaction
• Plus: high-quality generation
What's new:
• Fine-grained inline tags-steer [whisper], [angry], [breaths] & [laughs]
• Free-style natural-language control-“read this slowly, like a bedtime story”
• 16 languages
• Clean output even from noisy reference audio
• One-pass long-form up to 3 min
now #
1# on the Artificial Analysis TTS Leaderboard.
Blog:
API: