cs.AISep 29, 2026

HorizonFlow: Variable-Length Planning for Offline Goal-Conditioned RL

Authors: JunHyeok Oh, Zian Jang, Byung-Jun Lee

Organizations: Korea University. · Gauss Labs Inc., Seoul, Republic of Korea.

Abstract

Recent advances in generative planning have made trajectory inpainting a promising approach to offline goal-conditioned reinforcement learning. However, these methods typically specify the planning horizon before generating plan content, even though the appropriate horizon depends on the route itself. A horizon that is too short can force infeasible transitions, whereas one that is too long can introduce redundant motion. We introduce HorizonFlow, a hierarchical planner that treats plan length as an output of generation rather than a prescribed input. Its subgoal route planner guides its action-prefix controller through a sequence of latent subgoals. Both components combine insertion-based generation with flow matching to jointly generate continuous plan content and length, using the partially generated plan to guide token insertion. HorizonFlow reuses the resulting length information to select candidates and steer generation toward shorter plans without a separate learned value model. Across Maze2D, Multi2D, and OGBench navigation and visual manipulation benchmarks, HorizonFlow achieves the highest average performance among the compared methods.

Figures & tables

Appendix figures & tables20 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. HDFlow: Hierarchical Diffusion-Flow Planning for Long-horizon Tasks

    May 6, 2026Nandiraju Gireesh, Yuanliang Ju, Chaoyi Xu +3Diffusion PlanningClassical Planning

  2. Diffusion Subgoal Planning for Long-Horizon Offline Goal-Conditioned Reinforcement Learning

    Sep 28, 2026Hengrui Zhang, Yuhu Cheng, C. L. Philip Chen +1Goal-Conditioned Reinforcement LearningHigh-Level Subgoal Generation

  3. Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL

    Feb 3, 2026Jinwoo Choi, Sang-Hyun Lee, Seung-Woo SeoGoal-Conditioned Reinforcement LearningHigh-Level Subgoal Generation