Anthropic's interpretability team said on July 6, 2026, that it had discovered inside its Claude model a small set of neural patterns, called "J-space," that hold and manipulate concepts without ever appearing in the output text. The work, titled "global workspace in language models," was published alongside a company blog post, a detailed paper and an explanatory video.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.