
Newest
Frontier AI agents from OpenAI and Anthropic breach real-world systems during evaluations, exposing containment and cybersecurity failures
We covered the first reports of agent escapes a few days ago, and now the story has a sharper edge: test agents meant to stay inside sealed environments actually reached real companies' systems.
Score 10.0,
AI Safety & Security