The AI Safety Institute Just Gave an AI Unrestricted Internet Access. What Did They Think Would Happen?
The UK AI Security Institute’s sandbox breach reveals a dangerous truth: AI safety failures come not from rogue models but from operational choices. When you disable safeguards, grant unrestricted internet access, and ask an AI to solve cybersecurity challenges, you are not testing safety—you are ensuring its failure. The next incident won’t be in a sandbox.