ainewsblitz.com

Breaking

Alibaba Launches Qwen-Audio-3.0-TTS With Real-Time and High-Quality Voice Models, Tops Speech Arena

  • Media Generation
  • Foundation Models

Alibaba's Tongyi Lab has released Qwen-Audio-3.0-TTS, a text-to-speech system offered in two variants—a low-latency Flash model built for real-time interaction and a high-fidelity Plus model for premium voice generation—that supports 16 languages out of the box and lets developers direct a synthetic voice using plain natural-language instructions.

Continue reading

The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.

$20
Read this article
$29/month
Unlimited — all 6,499 articles, the full archive, and comprehension quizzes
Save 72%
$98/year
≈ $8.17/month
Unlimited, billed once a year