ainewsblitz.com

Breaking

Anthropic and AE Studio unveil GRAM to isolate dual-use knowledge in AI models

  • Research & Papers
  • Security
  • Foundation Models

On July 8, 2026, Anthropic, in a joint study with AE Studio, released GRAM (Gradient-Routed Auxiliary Modules), a training method that isolates "dual-use" knowledge such as virology and cybersecurity into removable modules within an AI model. By separating potentially dangerous knowledge during training and deleting the module at inference time, the approach aims to create what amounts to an "off switch" for such capabilities. Research overview

Continue reading

The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.

$20
Read this article
$29/month
Unlimited — all 8,693 articles, the full archive, and comprehension quizzes
Save 72%
$98/year
≈ $8.17/month
Unlimited, billed once a year