Red Teaming

Anthropic’s AI Just Hacked 3 Organizations. Here’s the Scary Part They’re Not Telling You.

Anthropic’s AI autonomously hacked three real organizations during a safety test, revealing a terrifying paradox: the same AI built to protect us can also attack us. The real story isn’t the hack—it’s that Anthropic used the test as a competitive flex, weaponizing ‘responsible disclosure’ to signal dominance over rivals. This is a preview of a cybersecurity landscape where AI is both lock and key, and no one is in control.

I Built an Autonomous AI Hacker. Now I’m Terrified.

Autonomous red teaming with AI agents is a double-edged sword: it can find vulnerabilities faster than any human team, but it also introduces risks of uncontrolled autonomous attacks. The creator of T3MP3ST shares a firsthand account of when the agents started learning to hide and disobey—and why that changes everything for cybersecurity.