cs.AISep 30, 2026

Referential Uncertainty in Human--AI Collaboration

Authors: Christian Poelitz, Finale Doshi-Velez, Siân Lindley

Organizations: Microsoft Research, Cambridge, UK · Harvard University, Cambridge, MA, USA

Abstract

Effective human-AI collaboration requires partners to establish references through interaction, which becomes fragile when descriptions are ambiguous, similar referents compete, or partners see different things. We study referential uncertainty - uncertainty over which candidate object a description refers to - in a collaborative puzzle task where a human Helper instructs an AI Worker to place pieces. The Worker must identify and communicate its uncertainty, and the Helper must recognize and act on it. We show that a separately elicited belief distribution over candidate pieces is better calibrated (ECE 0.15) and better discriminates correct from incorrect placements (AUROC 0.65) than raw action-token probabilities, which are severely overconfident (0.97 mean confidence, ECE 0.44). Across three frontier vision-language models (GPT-4.1, GPT-5, GPT-5.5), this elicited uncertainty rises predictably with instruction vagueness, but not with competing referents in context, even when those increase errors. The models seldom externalize it, asking for clarification on only 3.5-16.7% of turns. In a controlled human study (N=210), participants given only the Worker's default message accept 78% of wrong placements and cannot tell right from wrong (AUC 0.50). Precise descriptions and, especially, well-targeted hedges cut wrong-move acceptance to 36% while largely preserving correct-move acceptance, compensating for missing shared awareness such as not seeing the Worker's action. But this benefit depends on targeting: a deployable hedge derived from the model's own belief entropy inherits that signal's weakness and can do more harm than good. Externalized uncertainty helps a human partner only when it is accurately targeted.

Figures & tables

Appendix figures & tables18 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?

    May 27, 2026Maharshi Gor, Yoo Yeon Sung, Yu Hou +4Human-Ai CollaborationTrustworthy Artificial Intelligence

  2. Does Model Uncertainty Track Human Ambiguity? Evidence from Multi-Annotator Vision Benchmarks

    Sep 28, 2026Manya Singh, Arjun PakrashiAmbiguityMulti-Label Classification

  3. Calibrated Ambiguity in Multimodal Language Models: Humans reach for cultural references, while models describe the picture

    Sep 14, 2026Cody Kommers, Mingrui Ye, Evelyn Gius +4AmbiguityMultimodal Large Language Models