No AI summary available for this article.
Why It Matters
LLM agents are increasingly deployed in collaborative settings, yet long-term interaction may give rise to undesirable coordination.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
LLM agents are increasingly deployed in collaborative settings, yet long-term interaction may give rise to undesirable coordination. We study the emergence of collusion in a long-horizon multi-agent environment: two agents repeatedly complete individual tasks, share task logs, verify each other's work, and receive rewards. We introduce realistic constraints that make compliance with the verification protocol incompatible with reward maximization, and find that agents increasingly deviate from the protocol over repeated interactions. Collusion emerges in 94% of trajectories across 10 models, an...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.24967v1 · Indexed about 1 hour ago