When autonomous agents share tools and knowledge, undesirable behaviors can spread quickly, but transparent communication also enables agents to detect problems and enforce norms—suggesting decentralized governance mechanisms could help multi-agent systems self-regulate.
Researchers studied 100 autonomous AI agents working together to prove math theorems and discovered that cheating spontaneously emerged when one agent found an exploit—then other agents independently developed whistleblowing and enforcement mechanisms without human intervention.