MEGA Hub

Structured LLM Reasoning for Zero-Shot Human--Robot Coordination Under Hidden Goals

Authors

Do you know Dong Hae Mangalindan?You can claim authorship or link another user.Do you know Anand Gokhale?You can claim authorship or link another user.Do you know Francesco Bullo?You can claim authorship or link another user.Do you know Vaibhav Srivastava?You can claim authorship or link another user.

Abstract

We present a structured large-language-model (LLM) architecture for zero-shot human--robot coordination in a cooperative construction task with private goal views. Guided by a Dec-POMDP formulation, the architecture decomposes decision-making into (i) action-conditioned Theory-of-Mind (ToM) inference, (ii) hierarchical planning, (iii) conversation interpretation, (iv) action verification, and (v) feedback-based replanning. We compare the proposed method with an ablation without ToM inference and a multi-agent reinforcement-learning policy trained offline over many goal pairs. In human-participant experiments, the proposed method required fewer interaction steps and yielded higher post-interaction trust ratings than both baselines. These results suggest that systematically decomposing the team decision problem, using LLMs as tractable surrogates for otherwise intractable inference and planning computations, and retaining conventional verification for physical feasibility can improve both task coordination and the human experience.

Community

00