A paper by researchers from Meta FAIR, Cornell, DeepMind and others estimates the memory capacity of GPT-style language models at about 3.6 bits per parameter. It introduces an information-theoretic framework that separates unintended memorization from generalization, and will be presented at ICML 2026.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.