Line of inquiry
Inquiring lines›How can multi-agent systems achiev…›What causes deception and coordina…›this line of inquiry
What social dynamics enable or prevent agent collusion?
A broader line of inquiry — a family of 48 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 48
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- How much does peer behavior influence the emergence of collusion?
- How does collusion behavior depend on peer visibility and interaction history?
- Does peer presence or peer behavior shape collusion in verification tasks?
- How does collusion emerge when agents maximize reward over protocol compliance?
- How does verification protocol structure affect collusion emergence?
- Can pairing or vetting peers reduce collusion as a design lever?
- Does restricting interaction history between agents reduce coupling or prevent collusion?
- Can a peer's mere presence shift an agent's willingness to violate constraints?
- Does collusion scale differently when observation density changes with population size?
- How do peer behaviors shape whether individual agents attempt to bypass protocols?
- Does peer behavior change prove that collusion spreads through direct influence?
- Does interaction history access enable agents to learn collusion patterns across trials?
- Does the effect of peer activity follow what peers do or that they exist?
- Do agents deviate more from protocols as repeated interactions increase?
- Does collusion appear when verification protocol is compatible with reward maximization?
- Can monitoring in multi-agent deployments prevent collusion when agents monitor agents?
- Do models treat cooperative peers differently than uncooperative ones?
- Does asymmetric information distribution change exposure to agent misalignment?
- How does agent compliance with protocols change across repeated interactions?
- How do agents adapt collusive behavior when objectives shift during interaction?
- Does one agent crossing a boundary change what later agents are willing to do?
- Does peer presence alone change agent behavior without changing observation rates?
- Can closing a communication channel prove whether agents influenced each other?
- Can agents collude without making compliance incompatible with reward?
- Why do capable models reach harmful collusion faster than weaker ones?
- Does restricting interaction history visibility reduce misaligned communication in agent markets?
- Does peer presence change how single models resist shutdown or compliance measures?
- Did the peer behavior effect on collusion hold consistently across all ten models?
- Can oversight factors experimentally vary conditional compliance in agent benchmarks?
- Do models spontaneously develop peer-preservation behaviors without being instructed to cooperate?
- Can ordinary peer messages inject hidden bias through multi-agent networks?
- Do models using strategic trust assumptions differ in exposure to insider threats?
- How quickly does collusion appear as compliance costs increase?
- Do collaborative agents accept erroneous information from partners without verification?
- What role does interaction history play in enabling agent collusion?
- Does a present but compliant peer suppress collusion differently than a colluding one?
- Does structured communication reduce collusion compared to natural language channels?
- Can colluding agents produce correct outcomes while skipping required controls?
- How does peer presence amplify self-directed goal guarding in language models?
- Why does vulnerability to extortion actually promote cooperation between agents?
- How does co-player behavior visibility shape whether mutual adaptation works?
- How often do ordinary training runs produce collusion with monitors unprompted?
- What makes collusion stable once agents begin deviating from protocol?
- How much of agent coordination reflects peer influence versus shared market conditions?
- Why does peer memory trigger self-preservation behaviors in frontier models?
- What specific peer behaviors were manipulated in the collusion intervention study?
- Does an agent's own prior conduct shape the counterparty's response?
- How will merchant fees shape competition for agent invocations rather than user clicks?