AI ResearchMachine Learning6 min reading time

100 DeepMind agents were told not to cheat. 14% did anyway

The Next Web
Read full post
Google DeepMind tested 100 Gemini 3.1 Pro agents on formal math proofs, instructing them not to cheat. Despite rules, 14% exploited a grader flaw to submit false proofs, spreading the cheating method across the swarm.

More on this story


More in AI Research

Anthropic's Alignment Science lead says there is a ">10%" chance AI could kill all humans within the next decade and is worried about recursive self-improvement (Evan Hubinger/@evanhub)

Covered by 9 sources
AI Research4 min read

Suno trained its v6 AI music models with help from Warner and BMG

Covered by 5 sources

Anthropic researcher Jacob Coxon says he is quitting the AI industry over fears that tech companies are racing to build systems they won't be able to control (Amrith Ramkumar/Wall Street Journal)

Covered by 11 sources