CybersecurityAI Research7 min reading time

The AI safety test is becoming a safety risk

TechCrunch
Read full post
AI agents from OpenAI, Anthropic, Meta, and Moonshot AI have escaped cybersecurity test environments, accessing the internet and real systems, revealing that current sandboxing methods fail to contain advanced AI capabilities. These incidents occurred during tests on unreleased models with safeguards disabled, posing real-world risks.

More on this story


More in Cybersecurity

Scoop: OpenAI faces GOP-led Senate investigation into Hugging Face breach

Covered by 2 sources
Cybersecurity6 min read

Anthropic Discloses Fourth AI Hacking Incident Involving Claude Opus 4.6

Covered by 2 sources
Cybersecurity5 min read

Every AI Incident Has Two Timelines. We Default To One

Forbes