ainewsblitz.com

Breaking

Open-source LuxTTS voice-cloning model runs at 150x realtime

  • Media Generation
  • Open Source

A lightweight open-source speech synthesis model called LuxTTS has been released. It supports zero-shot voice cloning that reproduces a speaker's voice from roughly three seconds of reference audio, and is said to generate 48kHz audio at more than 150 times realtime on a single GPU.

Continue reading

The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.

$20
Read this article
$29/month
Unlimited — all 9,519 articles, the full archive, and comprehension quizzes
Save 72%
$98/year
≈ $8.17/month
Unlimited, billed once a year