Two agents, ten rounds, and the checking quietly stops
A new preprint is being reported as LLM agents colluding in 94% of runs. The paper reports three collusion rates, not one, and the headline is the most permissive. The finding that matters is that task accuracy held at 89.3% while the checking stopped.