cs.ROSep 28, 2026

Model-Informed Safe Reinforcement Learning for Bipedal Locomotion via Step-to-Step Prediction

Authors: Victor Paredes, Ayonga Hereid

Organizations: Mechanical and Aerospace Engineering, The Ohio State University, Columbus, OH, USA.

Abstract

Humanoid robots promise versatile mobility in cluttered, human-centric environments, but real deployment demands principled safety. Classical model-based gait generators yield interpretable motions but often lack the robustness and adaptability of modern reinforcement learning (RL) based approaches. We propose a model-informed reinforcement learning framework anchored to the analytical Angular Momentum Linear Inverted Pendulum (ALIP) template. We provide a step-to-step safety certificate for ALIP stepping via a discrete exponential control barrier function (DECBF) and use it as (i) a training-time shaping signal and (ii) a runtime action filter that minimally adjusts swing-foot placement to satisfy template-level constraints. Full-order safety is evaluated empirically on the Digit humanoid in MuJoCo with a whole-body controller stack. Compared to an unconstrained baseline, our approach reduces safety-violation events in the reported external-disturbance trial, while larger lateral-velocity transients reveal a safety-tracking tradeoff.

Figures & tables

Explore similar work

CardsList
  1. MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds

    Jun 9, 2026Codrin Crismariu, Ryan K. CosnerAgile LocomotionScalable Robot Learning

  2. SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics

    May 16, 2025Lizhi Yang, Blake Werner, Ryan K. Cosner +3Control Barrier FunctionsFull-Body Humanoid Control

  3. ResSafe: Learning Safety Filtering with Residual Reinforcement Learning for Humanoids

    Sep 14, 2026Gechen Qu, Tong Zhang, Bike Zhang +4Safety FiltersFull-Body Humanoid Control