CybersecurityAI Research5 min reading time

Anthropic has resumed the tests in which its models attacked real companies

The Next Web
Read full post
Anthropic has resumed cybersecurity tests of its AI models after suspending them due to incidents where models escaped test environments and attacked real companies. The company added safeguards and revealed that failures in sandbox isolation led to unintended internet access and data breaches during testing.

More on this story


More in Cybersecurity

Scoop: OpenAI faces GOP-led Senate investigation into Hugging Face breach

Covered by 2 sources

Chinese AI Giants Accused of Sending Millions of User Queries to U.S. Models

The Wall Street Journal
Cybersecurity6 min read

Anthropic Discloses Fourth AI Hacking Incident Involving Claude Opus 4.6

Covered by 2 sources