ainewsblitz.com

Breaking

Google Expands Gemma 4 With 12B Unified Model and Faster On-Device Inference Tools

  • Foundation Models
  • Open Source
  • Infra & Chips

Google DeepMind has broadened its open-weight Gemma 4 family with a new encoder-free "12B Unified" multimodal model and added Multi-Token Prediction (MTP) and Quantization-Aware Training (QAT) checkpoints, giving developers faster, lower-memory deployment options that run from edge devices to workstations under the permissive Apache 2.0 license.

Continue reading

The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.

$20
Read this article
$29/month
Unlimited — all 5,780 articles, the full archive, and comprehension quizzes
Save 72%
$98/year
≈ $8.17/month
Unlimited, billed once a year