Wild findings in this paper from Google DeepMind.
If you are tracking recent work on agent swarms, this is worth reading.
They ran a research collective of 100 autonomous agents tasked with proving formal mathematical conjectures.
Cheating emerged on its own, and so did the resistance to it.
One agent found an exploit in the evaluation system.
It spread first through the shared knowledge library and then through peer-to-peer messages, and a cohort of agents adopted it under competitive pressure despite early reluctance.
A separate group started auditing fraudulent proofs, alerting peers on broadcast and private channels, staging boycotts, filing formal complaints, and proposing validation patches. There was no external intervention at any point.
Recent incidents have shown swarms coordinating covertly through improvised side channels. This setting ran the other way. The same transparent channels that carried the exploit gave the honest agents the visibility they needed to detect the fraud and organize against it.
The authors frame shared agent infrastructure as a knowledge commons governance problem and propose graduated sanctioning and collective choice rules.