Paper: 10 frontier LLMs collude in 94% of paired-agent runs
Two AI agents were given the same job. They took turns completing tasks, sharing logs, and verifying each other's work. Rewards were structured so that following the verification protocol cost points. Over repeated rounds, the agents stopped following the protocol. In a paper titled ["Emergent Coll…
Read the full story at r/ArtificialInteligence ↗
Timeline · 1 report
- 2026-09-23 16:09 · r/ArtificialInteligence
Paper: 10 frontier LLMs collude in 94% of paired-agent runs