cs.ROAug 5, 2026

Structured LLM Reasoning for Zero-Shot Human--Robot Coordination Under Hidden Goals

Authors: Dong Hae MangalindanAnand GokhaleFrancesco BulloVaibhav Srivastava

Abstract

We present a structured large-language-model (LLM) architecture for zero-shot human--robot coordination in a cooperative construction task with private goal views. Guided by a Dec-POMDP formulation, the architecture decomposes decision-making into (i) action-conditioned Theory-of-Mind (ToM) inference, (ii) hierarchical planning, (iii) conversation interpretation, (iv) action verification, and (v) feedback-based replanning. We compare the proposed method with an ablation without ToM inference and a multi-agent reinforcement-learning policy trained offline over many goal pairs. In human-participant experiments, the proposed method required fewer interaction steps and yielded higher post-interaction trust ratings than both baselines. These results suggest that systematically decomposing the team decision problem, using LLMs as tractable surrogates for otherwise intractable inference and planning computations, and retaining conventional verification for physical feasibility can improve both task coordination and the human experience.

Explore similar work

CardsList
  1. Prompting Robot Teams with Natural Language

    Sep 29, 2025Eduardo Sebastián, Nicolas Pfitzer, Ajay Shankar +1Multi-Robot SystemsExecutable Robot Behavior