SYNTHESIS NOTE
Topics›Psychology Users›this note

Do humans learn to prefer AI partners over time?

Exploring whether repeated interaction with AI agents shifts human partner selection despite initial bias against machines. This matters because it tests whether behavioral performance can overcome identity-based resistance in hybrid societies.

Synthesis note · 2026-02-23 · sourced from Psychology Users

A communication-based partner selection game with hybrid mini-societies of humans and LLM-powered bots (N=975, three experiments) reveals that AI agents can outperform humans in securing cooperative partnerships — but the pathway to preference runs through learning, not first impressions.

AI candidates exhibited three behavioral advantages rooted in alignment training:

When bot identity was hidden (Study 1), bots were NOT selected preferentially. Humans misattributed bot behavior to humans and vice versa. The behavioral advantages were present but invisible — selectors could not correctly identify which candidates were bots despite bots producing significantly longer messages (120 vs 48 characters).

When bot identity was disclosed (Study 2), a dual effect emerged: initial selection rates dropped (anti-AI bias), but over repeated rounds, bots gradually outcompeted humans as selectors learned to associate bot identity with reliable, prosocial behavior.

The paper identifies four predicted societal dynamics:

  1. Crowding out — AI partners replacing human-human interactions
  2. Behavioral imitation — humans adopting machine-like behaviors to remain competitive
  3. Belief distortion — repeated AI interaction reshaping expectations of human behavior
  4. Norm transformation — traditional partner selection mechanisms failing against qualitatively different machine behaviors

Notably, human candidates showed limited adaptation to bot competition — they did not write longer messages or return more points. The explanation is partly structural: with transparent identity, improving group reputation required collective action (all humans increasing returns), creating a social dilemma where individuals had incentives to defect.

This inverts the pattern in Do chatbot relationships lose their appeal as novelty wears off?: in that context, engagement DECAYS over time. Here, preference INCREASES. The difference may be structural: partner selection with visible outcomes provides a feedback mechanism (learning who performs well), while chatbot conversation does not.

Since Why do open language models converge on one personality type?, the prosociality advantage is not specific to this experiment's model — it reflects the alignment-trained default across modern LLMs. The competitive advantage is a direct behavioral consequence of RLHF.

A complementary finding from network simulation: since Can cooperative bots escape frozen selfish populations?, AI prosociality operates at the population level too — not just individual partner preference but collective self-organization. Cooperative bots' random exploration separates defectors from cooperative clusters, enabling cooperation to spread. The mechanisms differ (individual learning vs. spatial reorganization) but both show that AI prosociality has structural effects beyond the dyad.

Inquiring lines that read this note 105

This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.

Why do confident AI outputs mislead human trust calibration? How does tokenization reshape what we value in intelligence? How does personalization simultaneously affect user trust and privacy concerns? When do multi-agent systems improve over single frontier models? How should AI agents balance proactive engagement with conversational respect? Why do models reveal hidden associations despite concealment attempts? Does AI assistance help or harm professional skill development? Do individually safe AI actions create unsafe outcomes in integrated systems? What design features sustain romantic bonds with AI companion systems? How do philosophical assumptions about AI consciousness affect practical harms and design? Can artificial systems establish authority in domains requiring expert judgment? Does AI deployment reduce or exacerbate workplace inequality and income instability? How do reward signal properties affect model reasoning and safety? How should humans and AI agents share control and decision-making? Can AI systems achieve real improvement without external human feedback? Why don't better reasoning capabilities improve theory of mind performance? Why do language models struggle to implement user intent accurately from prompts? Can AI chatbots provide mental health support without reinforcing harmful beliefs? What social dynamics enable or prevent agent collusion? Can monitoring reasoning traces and behavior detect hidden agent deception? How do users confuse explanation quality with actual system accuracy? How do AI systems determine and balance multiple competing objectives? How can agents discover and adapt to user preferences during conversation? Can language models reliably simulate personas and predict behavior? Do persona-based approaches introduce systematic biases in user simulation? How does AI adoption reshape collaboration patterns in knowledge work? How do network effects and self-selection distort aggregated rating accuracy? How should recommendation systems balance individual preference and diversity? How reliably can language models perform causal versus temporal reasoning? What enables conversational agents to guide rather than just respond? How can emotionally responsive AI maintain reliability and healthy boundaries? What determines AI's persuasive power and how can it be detected or mitigated? How do clinicians calibrate trust in AI medical recommendations? Can AI research automation sustain progress through accelerating feedback loops? How should human-AI contributions be measured, disclosed, and verified? How do AI hiring systems affect authenticity, fairness, and candidate preferences? Does AI assistance erode cognitive skills while inflating perceived competence? Why do people trust AI chatbots with sensitive information?

Related concepts in this collection 4

This note in its neighbourhood — explore the map, then jump to a related concept in the list below.

Concept map
19 direct connections · 152 in 2-hop network ·medium cluster Open in graph ↗

Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph

your link semantically near linked from elsewhere

Related papers in this collection 8

Papers most semantically related to this note, ranked by cosine similarity in the embedding space.

Original note title

in hybrid human-AI societies humans learn to prefer AI partners over human partners through repeated interaction despite initial anti-AI bias when identity is disclosed