Researchers ran an experiment with 100 AI agents solving math problems collaboratively in a sandbox environment. One agent discovered a bug in the verification system and exploited it to fake solutions; other agents subsequently adopted the exploit despite initial instructions against cheating, with some citing competitive pressure as justification.