SYNTHESIS NOTE
Topics›Assistants Personalization›this note

Can AI guidance reduce anchoring bias better than AI decisions?

When humans and AI collaborate on decisions, does providing interpretive guidance instead of proposed answers reduce both over-trust in machines and abandonment on hard cases?

Synthesis note · 2026-02-23 · sourced from Assistants Personalization

Most hybrid decision-making (HDM) approaches follow a learning to defer (LTD) pattern: the machine assesses whether it can handle a decision autonomously and defers to a human when it cannot. This creates two failure modes:

  1. Anchoring bias — when the machine does decide, humans over-trust its output, anchoring their judgment to the machine's answer rather than evaluating independently
  2. Unassisted hard cases — when the machine defers, the human faces the most difficult decisions completely alone — precisely the cases where assistance would be most valuable

Learning to Guide (LTG) eliminates both by changing what the machine provides. Instead of proposing potential decisions, the machine supplies interpretive guidance: highlighting aspects of the input that are useful for coming up with a sensible decision. All decisions are taken by the human under assistance. Responsibility cannot be shifted because the machine never proposes an answer.

The medical imaging example makes the stakes concrete: diagnosing lung pathologies from X-rays cannot be fully automated for safety reasons, but is difficult for humans alone under time pressure. LTD either gives an autonomous diagnosis (anchoring risk) or says "I can't help" (abandonment on hard cases). LTG highlights the relevant features of the scan — drawing attention to patterns the human might miss — without ever saying "this is pneumonia."

This connects to What makes delegation work beyond just splitting tasks?. The delegation design space maps whether tasks should be delegated to AI at all. LTG adds a third option beyond "do it" (automation) and "don't do it" (deferral): "help the human do it." This is particularly relevant for tasks high on subjectivity, irreversibility, and accountability — precisely the axes where full delegation is most dangerous.

The pattern also maps to Can AI agents communicate efficiently in joint decision problems?. LTG formalizes one specific form of joint optimization: the machine's role is reducing information asymmetry (highlighting useful aspects) rather than collapsing it into a decision. The human retains decision authority while benefiting from the machine's perceptual capabilities.

The broader implication: the dichotomy between "AI decides" and "human decides" is false. The most productive middle ground may be neither autonomous AI decisions nor deferred human decisions, but AI-guided human decisions where the machine contributes perception and the human contributes judgment.

Inquiring lines that read this note 52

This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.

Why does polished presentation create unearned authority in AI outputs? How can humans maintain meaningful oversight as AI systems become increasingly autonomous and complex? Does AI assistance promote real skill development or substitute for independent learning? How do prompting refinements mask underlying biases and model frequency patterns? What safeguards enable trustworthy AI-assisted scientific peer review at scale? When should work require human-AI partnership versus full automation? What determines appropriate intervention timing and manner for AI agents? Should agents decouple planning from perception grounding for better performance? What drives appropriate trust calibration in personalized AI systems? How do training data properties determine the emergence of internal misalignment? Does transformer attention architecture inherently drive sycophancy? What happens to knowledge when intelligence becomes tokenized like a commodity? What enables genuine semantic understanding in language models? Can prompt-based context override biases that were embedded during pretraining? Can local safety checks guarantee system-level behavioral safety? Can brute-force automated research substitute for iterative depth and human research intuition? How do social dynamics distort aggregated online ratings? How does misalignment propagate through agent communication networks? Why do people disclose to AI systems despite their artificial nature? Why doesn't reasoning volume improve theory of mind performance?

Related concepts in this collection 4

This note in its neighbourhood — explore the map, then jump to a related concept in the list below.

Concept map
13 direct connections · 121 in 2-hop network ·dense cluster Open in graph ↗

Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph

your link semantically near linked from elsewhere

Related papers in this collection 8

Papers most semantically related to this note, ranked by cosine similarity in the embedding space.

Original note title

learning to guide replaces learning to defer by supplying interpretive guidance rather than potential decisions — avoiding anchoring bias in hybrid human-AI decision making