cs.ROOct 5, 2026

Task-Space Imitation Guidance for Efficient Reinforcement Learning

Authors: Salar Asayesh, Hossein Darani, Todd Cao, Evgeny Andriash, Mani Ranjbar

Organizations: Sanctuary AI, Canada

Abstract

We introduce Task-Space Imitation Guidance for Efficient Reinforcement Learning (TIGER), a reward-construction and pretraining framework for sparse-reward tabletop robotic manipulation. TIGER treats an action-chunked imitation policy not as an executable controller or action prior, but as a local task-space progress estimator: predicted action chunks are converted, using controller-aware action-to-motion mapping, into short-horizon end-effector references, and the RL agent receives dense progress rewards toward these references while the sparse environment reward remains the dominant objective. During pretraining, TIGER uses imitation-guided look-ahead signals to relax conservative value penalties for actions predicted to make task-space progress, reducing off-manifold exploration during early online RL. Across simulation and real-robot experiments, TIGER improves early sample efficiency and reduces measured safety violations while matching or improving final success rates relative to prior RL and IL-RL baselines on the evaluated tasks.

Figures & tables

Appendix figures & tables14 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

    Jun 8, 2026Yuying Zhang, Francesco Verdoja, Wenyan Yang +1Robot Skill LearningRobotic RL