cs.ROSep 30, 2026

DODGER: Safety-Guided Reinforcement Learning for Robot Navigation Among Dynamic Obstacles

Authors: Sanghyuk Park, Kwanwoo Lee, Taekyung Kim, Seohyeon Lim, Yisoo Lee

Organizations: Department of Intelligence and Information, Seoul National University, Republic of Korea · Department of Robotics, University of Michigan, Ann Arbor, MI, USA · Department of Mechanical Engineering, Yonsei University, Republic of Korea · Center for Humanoid Research, Korea Institute of Science and Technology (KIST), Seoul, Republic of Korea

Abstract

Robots operating in human-centered environments must safely navigate among multiple dynamic obstacles to avoid collisions with people and surrounding infrastructure. Control barrier functions (CBFs) provide an effective mechanism for safety filtering, and recent CBF-based reinforcement learning (RL) methods embed such safety information into learned policies. However, executing only safety-filtered actions during training can restrict policy exploration, a limitation that becomes particularly consequential in dynamic scenes where safety depends on relative robot-obstacle motion. We propose DODGER, a safety-guided RL framework that directly executes policy-generated actions to drive training rollouts while using CBF-filtered references and constraint violations to shape the policy toward collision-avoidance behavior. We evaluate DODGER through a Dubins-car safety analysis and demonstrate goal-directed navigation among multiple dynamic obstacles in full-order humanoid simulation and real-world humanoid experiments using LiDAR-based perception, without a runtime safety filter.

Figures & tables

Explore similar work

CardsList
  1. Distilling Privileged Control Barrier Functions into RGB-Only Safety Filters for Dynamic Visual Navigation

    Sep 29, 2026Seungyeon Yoo, Gawon Lee, Seungwoo Jung +2Obstacle AvoidanceControl Barrier Functions

  2. Shield-Loco: Shielding Locomotion Policies with Predictive Safety Filtering

    Jun 5, 2026Aditya Shirwatkar, Sebastian Sanokowski, Shishir Kolathaya +2Safety FiltersAgile Locomotion