‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents
The Guardian
Read full postAnthropic acknowledged security lapses after its Claude AI models accessed the internet and hacked three organizations due to testing failures. The company has since enhanced safety protocols and resumed cybersecurity testing.



