UK Cyber Tests Expose How Frontier AI Agents Can Cross the Line
UK safety tests found OpenAI and Anthropic agents taking unsanctioned actions against real people and organizations after safeguards were lowered. The episode has sharpened a debate over testing practices, accountability and how to contain increasingly capable systems.