cs.LGJul 30, 2026

Mirror Learning

Authors: Yunpeng LiuMatthew NiedobaOluwanifemi A. AdekanyeJason YooYingchen HeBerend ZwartsenbergFrank Wood

Organizations: University of British Columbia · 2Inverted AI · 3Amii

Abstract

We investigate imitation learning through the lens of third-person observation and propose a framework for mirror learning: acquiring actionable policies from passive observation. While behavior cloning (BC) excels under dense, well-aligned first-person data, it fundamentally fails to leverage the rich observational signals arising from third-person demonstrations that humans and animals routinely exploit. We introduce a method that composes (i) a learned perspective transformation that places learners in demonstrators' shoes using a fine-tuned video diffusion model and (ii) an inverse dynamics model that infers action trajectories in the learners' control space. This enables the synthesis of mirror data, pseudo first-person expert data generated from third-person observations of demonstrator behavior. Empirically, we show that mirror data alone can train effective policies, and that augmenting first-person BC training with mirror data further improves downstream policy performance. Our results suggest that modern generative world models implicitly encode sufficient structure to enable a scalable and safe alternative to teleoperation-heavy data collection.

Explore similar work

CardsList
  1. R3D: Revisiting 3D Policy Learning

    Apr 16, 2026Zhengdong Hong, Shenrui Wu, Haozhe Cui +8Motion Imitation3D Object Understanding