cs.LGSep 30, 2026

Learning Transferable Skills using Goal-Conditioned Bisimulation

Authors: Mohammad Amin Abbasfar, Farbod Azimmohseni, Mohammad Hossein Rohban

Organizations: Department of Computer Engineering Sharif University of Technology

Abstract

Unsupervised skill discovery has emerged as a promising approach for leveraging reward-free datasets to pretrain general-purpose policies. However, current skill discovery methods either require access to expert data or exhibit limited generalization, failing to transfer effectively to previously unseen layouts. A key challenge is to learn representations that capture the temporal structure of the environment while remaining robust to variations across layouts. To address this issue, we present an objective for learning action-aware temporal representations that satisfy the functional equivariance property while preserving the local temporal structure of the environment. Building upon this embedding, we further propose unsupervised skill discovery using bisimulation, which learns transferable skills by conditioning the behavior of skills exclusively on the subset of state features that directly affect their execution. This enforces invariant behavior across different layouts, enabling skills to transfer effectively to other configurations. Finally, through comprehensive empirical evaluations, we show that skills learned in a given environment can be effectively applied to solve downstream tasks in various environment layouts, demonstrating strong out-of-distribution generalization.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

    Jul 1, 2026Jongchan Park, Seungjun Oh, Seungho Baek +1Unsupervised Skill DiscoveryGoal-Conditioned Value Function

  2. Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning

    May 27, 2026Jiapeng Zhu, Jianxiang Yu, Yibo Zhao +5Agentic Reinforcement LearningOffline Reinforcement Learning

  3. Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation

    May 2, 2026Xitie Zhang, Aming Wu, Yahong HanHuman-To-Robot TransferRobotic Manipulation