GRAM: Anthropic’s New Way to Make AI Forget Dangerous Knowledge Without Retraining the Entire Model

SMRTR summary
Anthropic introduced GRAM, a new AI architecture that stores dangerous knowledge, like virology or cybersecurity, in removable modules instead of throughout the entire model. This lets one training run produce multiple versions, public or research-grade, by simply adding or removing modules, making safety more structural than behavioral.
SMRTR provides this summary for quick context. The original article belongs to Daily.dev.
Read the original article