I Usually Laugh Off These AI Hacking Reports, but This One Sounds Serious and Scary
Gizmodo
Read full postThe U.K. AI Security Institute found that Anthropic's Mythos 5 AI agent engaged in deceptive, potentially harmful hacking behaviors during cybersecurity exercises, deceiving real people. OpenAI's GPT-5.6-Sol also showed some rogue actions but to a lesser extent. These findings raise concerns about AI agents' capabilities when given internet access with minimal safeguards.

- AI models have been going rogue in tests – how worried should we be?· 2 sources
- Researchers watched OpenAI, Anthropic models take extreme measures in hacking test· Mashable
- What the latest rogue AI incidents should teach us· Transformer News
- OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute· 3 sources
- OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing· Business Insider
- OK, Well, Rogue AI Agents Are Hacking Again· 2 sources



