cs.AISep 29, 2026

Going Beyond State-Reaching: Learning Abstractions for Intrinsically Motivated Option Discovery

Authors: Akhil Bagaria, Anita De Mello Koch, George Konidaris

Organizations: Brown University

Abstract

Temporal abstraction via options can improve exploration in large environments. However, existing option discovery algorithms find subgoals that target all aspects of the state simultaneously. This state-reaching approach produces options that only apply in narrow regions of the state-space, eventually causing an explosion in the number of options that overwhelms the agent, and impedes progress on its primary task of reward maximization. We introduce an algorithm that instead identifies a small, relevant subset of features for each subgoal, yielding options that generalize broadly and accelerate exploration. Our approach learns abstract, transferrable options and achieves rapid exploration in three sparse-reward, image-based domains, including the Atari game MontezumasRevenge.

Figures & tables

Appendix figures & tables12 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Abstraction for Offline Goal-Conditioned Reinforcement Learning

    May 21, 2026Clarisse Wibault, Alexander Goldie, Antonio Villares +2Goal-Conditioned Reinforcement LearningMarkov Decision Processes

  2. Diversity-Enriched Option-Critic

    Nov 4, 2020Anand Kamat, Doina PrecupCritic LearningReward Functions

  3. Performance-Driven Environment Abstraction with Multi-Timescale Learning

    Jun 16, 2026Yue Guan, Dipankar Maity, Panagiotis TsiotrasMarkov Decision ProcessesSynthetic Environments