On July 1, 2026, NVIDIA Research released Nemotron-Labs-TwoTower, a diffusion language model that splits an existing 30B-scale model into two towers to generate tokens in parallel. The open-weight model reportedly retains 98.7% of the base model's quality while boosting generation throughput by 2.42x.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.