-+ 0.00%
-+ 0.00%
-+ 0.00%

Qianwen announced: Today, we are officially releasing Qwen-Audio-3.0-TTS, a large speech synthesis model, including a Flash version for real-time interaction and a Plus version for high-quality generation. The new model achieves systematic improvements in fine-grained label control, freestyle instruction compliance, multi-language and dialect coverage, and complex acoustic robustness, moving synthesized speech from “being able to speak” to “able to express”.

Zhitongcaijing·07/20/2026 08:57:05
Listen to the news
Qianwen announced: Today, we are officially releasing Qwen-Audio-3.0-TTS, a large speech synthesis model, including a Flash version for real-time interaction and a Plus version for high-quality generation. The new model achieves systematic improvements in fine-grained label control, freestyle instruction compliance, multi-language and dialect coverage, and complex acoustic robustness, moving synthesized speech from “being able to speak” to “able to express”.