Alibaba's video-generation team Wan has unveiled Wan-Streamer, a full-duplex conversation model that handles audio, video and text in a single model, listening while responding with synchronized speech and video. The team published a v0.1 paper in June 2026 and a v0.2 paper in early July, with previews shown on the official site and a Hugging Face collection.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.