The AI Industry’s Darkest Secret: They’re Teaching Robots to Forget
Anthropic’s research proposes a method to selectively remove dangerous dual-use knowledge from AI models, forcing an uncomfortable choice: keep AI smart and risk catastrophe, or lobotomize it and lose life-saving breakthroughs. The safest AI might be the one that knows less — but that’s a terrifying future.