AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
Summary
This paper studies how humans decide when to delegate to AI and when to adopt AI suggestions in cooperative question answering, finding that confirmation bias drives suboptimal trust decisions such as under-reliance on correct AI outputs.
View Cached Full Text
Cached at: 06/02/26, 03:35 PM
Paper page - AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
Source: https://huggingface.co/papers/2605.28255
Abstract
Human-AI collaboration in question-answering tasks reveals suboptimal reliance decisions where humans under-rely on correct AI suggestions and over-rely when AI misleads them, with confirmation bias contributing to reduced trust in conflicting AI outputs.
AI systems are fallible, and humans can make mistakes in deciding whether totrustAI over their own judgment. Thus, improvinghuman-AI collaborationrequires understanding when, why, and how humans decide to rely on AI. We study two distinct reliance decisions: thedelegation choice-- deciding when to let AI act autonomously without knowing its output, and theadoption choice-- evaluating AI suggestions and deciding how to use them. Both of these decoupled reliance patterns shape collaboration, but prior work rarely studies them together in realistic settings with the same users. We address this gap by studying collaborative human--AI teams competing in aquestion-answering gamein which humans can choose when and how to work withAI agentsto win. Our 24 matches pair 23 expert humans with 16AI agents, capturing 387 delegation and 1440 adoption decisions. While human--AI collaboration performs better than either AI or humans alone, humans make suboptimal collaboration decisions, both under-relying on correct AI suggestions (3.9% of opportunities missed) and over-relying when AI misleads them (1.7%). Both parties contribute wrong answers: reported model confidence is near chance when humans and AI disagree, whileconfirmation biasdrives higher under-reliance (64.5%) when an AI suggestion agrees with humans’ initial incorrect answer. To close this gap, we recommendcalibrated confidence,evidence-grounded explanations, and mechanisms that help users refinetrust.
View arXiv pageView PDFAdd to collection
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2605.28255 in a model README.md to link it from this page.
Datasets citing this paper0
No dataset linking this paper
Cite arxiv.org/abs/2605.28255 in a dataset README.md to link it from this page.
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2605.28255 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
What would actually make you trust an AI? Not "it sounds right," but trust it the way you trust a person or an institution?
A discussion exploring what specific conditions (transparency, verifiable track record, persistent identity, accountability) would make people trust AI systems as they trust humans or institutions, rather than just accepting them as tools.
(Human) Attention Is (Still) All You Need: Human oversight makes AI-assisted social science reliable
This paper proposes that reliability in AI-assisted social science research depends on decision architecture—how cognitive labor is divided between humans and machines. Through a pre-specified factorial experiment, the authors show that an unconstrained multi-agent baseline fails in 72% of runs, while one organized with three architectural commitments (LLMs restricted to reasoning, deterministic data/estimation, and three human decision gates) fails in only 16%.
The Trust–Oversight Paradox: As AI Gets Better, Humans May Stop Really Overseeing It
A thought piece arguing that as AI becomes more accurate, human oversight may degrade into routine approval, creating a 'Trust–Oversight Paradox' where high-performing AI can still fail due to incomplete representation, stale data, or automation bias, suggesting a shift from human review to governing boundaries.
At what point would you trust an AI agent more than a new employee?
A discussion on the threshold for trusting AI agents versus new human employees, weighing tasks like lead qualification and scheduling against human-only roles like customer escalations and contract negotiations.
AI vs humans , whom do you trust more in 2026
A discussion on whether people find it easier to discuss personal topics with AI or humans, noting that AI offers a non-judgmental, always-available ear but lacks genuine human experience.