cs.CVSep 29, 2026

Multi-Granularity Language-Guided Imitation Learning via Instruction Decomposition

Authors: Yi-Pei Chiu, Wei-Ta Chu

Organizations: National Cheng Kung University, Taiwan

Abstract

Using language instructions as conditions to guide robot policy learning has recently become an important research domain. However, existing language-guided policy learning methods typically use an overall task description to guide the entire demonstration trajectory. For manipulation tasks involving multiple execution stages, these methods assign the same language description to different subtasks, making it difficult to distinguish the behaviors required at different stages. In this work, we propose a multi-granularity language guidance method based on instruction decomposition. The proposed method decomposes an overall task description into more fine-grained, concrete subtask-level language instructions, thereby enhancing learning efficiency and improving performance. We evaluate the proposed method in the setting of multi-task imitation learning and validate its effectiveness.

Figures & tables

Explore similar work

CardsList
  1. Vision-Language Grounded Task-Context-Aware Imitation Learning for Robotic Disassembly

    Sep 15, 2026Jeon Ho Kang, Igal Tamarkin, Ethan Niu +2Robotic ManipulationStrong Imitation Learning

  2. Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control

    Date pendingJun Chen, Erdemt Bao, Wenlong Dong +7

  3. Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents

    Apr 18, 2026Sukai Huang, Chenyuan Zhang, Fucai Ke +4Fine-Grained ActionsEmbodied Agents