Alibaba's Tongyi Lab has released Qwen-Audio-3.0-TTS, a text-to-speech system offered in two variants—a low-latency Flash model built for real-time interaction and a high-fidelity Plus model for premium voice generation—that supports 16 languages out of the box and lets developers direct a synthetic voice using plain natural-language instructions.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.