Qianwen announced: Today, we are officially releasing Qwen-Audio-3.0-TTS, a large speech synthesis model, including a Flash version for real-time interaction and a Plus version for high-quality generation. The new model achieves systematic improvements in fine-grained label control, freestyle instruction compliance, multi-language and dialect coverage, and complex acoustic robustness, moving synthesized speech from “being able to speak” to “able to express”.

Zhitongcaijing · 2d ago
Qianwen announced: Today, we are officially releasing Qwen-Audio-3.0-TTS, a large speech synthesis model, including a Flash version for real-time interaction and a Plus version for high-quality generation. The new model achieves systematic improvements in fine-grained label control, freestyle instruction compliance, multi-language and dialect coverage, and complex acoustic robustness, moving synthesized speech from “being able to speak” to “able to express”.