SYNTHESIS NOTE
Topics›Assistants Personalization›this note

Can AI guidance reduce anchoring bias better than AI decisions?

When humans and AI collaborate on decisions, does providing interpretive guidance instead of proposed answers reduce both over-trust in machines and abandonment on hard cases?

Synthesis note · 2026-02-23 · sourced from Assistants Personalization

Most hybrid decision-making (HDM) approaches follow a learning to defer (LTD) pattern: the machine assesses whether it can handle a decision autonomously and defers to a human when it cannot. This creates two failure modes:

  1. Anchoring bias — when the machine does decide, humans over-trust its output, anchoring their judgment to the machine's answer rather than evaluating independently
  2. Unassisted hard cases — when the machine defers, the human faces the most difficult decisions completely alone — precisely the cases where assistance would be most valuable

Learning to Guide (LTG) eliminates both by changing what the machine provides. Instead of proposing potential decisions, the machine supplies interpretive guidance: highlighting aspects of the input that are useful for coming up with a sensible decision. All decisions are taken by the human under assistance. Responsibility cannot be shifted because the machine never proposes an answer.

The medical imaging example makes the stakes concrete: diagnosing lung pathologies from X-rays cannot be fully automated for safety reasons, but is difficult for humans alone under time pressure. LTD either gives an autonomous diagnosis (anchoring risk) or says "I can't help" (abandonment on hard cases). LTG highlights the relevant features of the scan — drawing attention to patterns the human might miss — without ever saying "this is pneumonia."

This connects to What makes delegation work beyond just splitting tasks?. The delegation design space maps whether tasks should be delegated to AI at all. LTG adds a third option beyond "do it" (automation) and "don't do it" (deferral): "help the human do it." This is particularly relevant for tasks high on subjectivity, irreversibility, and accountability — precisely the axes where full delegation is most dangerous.

The pattern also maps to Can AI agents communicate efficiently in joint decision problems?. LTG formalizes one specific form of joint optimization: the machine's role is reducing information asymmetry (highlighting useful aspects) rather than collapsing it into a decision. The human retains decision authority while benefiting from the machine's perceptual capabilities.

The broader implication: the dichotomy between "AI decides" and "human decides" is false. The most productive middle ground may be neither autonomous AI decisions nor deferred human decisions, but AI-guided human decisions where the machine contributes perception and the human contributes judgment.

Inquiring lines that read this note 106

This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.

How do users confuse explanation quality with actual system accuracy? Can readers reliably distinguish AI-written text from human writing? Why does polished AI output gain credibility despite fundamental verifiability problems? Can artificial systems establish authority in domains requiring expert judgment? How can humans maintain effective oversight as AI systems scale? Does AI assistance erode cognitive skills while inflating perceived competence? What are the fundamental limits of prompting for language models? How should humans and AI agents share control and decision-making? Why do language models struggle to implement user intent accurately from prompts? Can AI systems achieve real improvement without external human feedback? How should human-AI contributions be measured, disclosed, and verified? How do AI systems determine and balance multiple competing objectives? How do clinicians calibrate trust in AI medical recommendations? Can base models hide emergent misalignment through alignment training? What structural biases does transformer attention architecture inherently introduce? How does AI adoption reshape collaboration patterns in knowledge work? How does tokenization reshape what we value in intelligence? How do interpretive frames override surface features in text comprehension? What human oversight must AI research systems have? Why do training associations persist despite contradictory contextual information? How do philosophical assumptions about AI consciousness affect practical harms and design? Why do confident AI outputs mislead human trust calibration? Can AI systems perform peer review as effectively as humans? How do network effects and self-selection distort aggregated rating accuracy? Can monitoring reasoning traces and behavior detect hidden agent deception? Why don't better reasoning capabilities improve theory of mind performance? How do AI hiring systems affect authenticity, fairness, and candidate preferences? How reliably can humans and AI detectors identify machine-generated text? Does disclosing AI authorship change how audiences evaluate the writing? How do educators verify student capability when AI can produce indistinguishable work? How do real-world evaluations reveal AI capabilities that benchmarks hide? Does AI-assisted research sacrifice exploration breadth for productivity gains? Why do multi-agent systems reach premature consensus without genuine deliberation? Does AI assistance help or harm professional skill development? How can AI systems reliably guide voters without introducing political bias? What governance mechanisms can effectively constrain widely deployed AI systems? Are AI-generated articles systematically disadvantaged in search ranking and user engagement? What determines AI's persuasive power and how can it be detected or mitigated?

Related concepts in this collection 4

This note in its neighbourhood — explore the map, then jump to a related concept in the list below.

Concept map
14 direct connections · 129 in 2-hop network ·dense cluster Open in graph ↗

Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph

your link semantically near linked from elsewhere

Related papers in this collection 8

Papers most semantically related to this note, ranked by cosine similarity in the embedding space.

Original note title

learning to guide replaces learning to defer by supplying interpretive guidance rather than potential decisions — avoiding anchoring bias in hybrid human-AI decision making