ainewsblitz.com

Breaking

Microsoft Open-Sources VibeVoice-ASR, a Long-Form Speech Recognition Model That Runs Locally Across 50+ Languages

  • Open Source
  • Foundation Models
  • Media Generation

Microsoft has released VibeVoice-ASR, an open-source speech recognition model built for long-form audio that runs entirely on a local machine, supports more than 50 languages, and removes the need for paid transcription APIs. The model is available under the permissive MIT license through the company's VibeVoice repository on GitHub and as a downloadable model card on Hugging Face.

Continue reading

The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.

$20
Read this article
$29/month
Unlimited — all 7,041 articles, the full archive, and comprehension quizzes
Save 72%
$98/year
≈ $8.17/month
Unlimited, billed once a year