cs.ROOct 7, 2026

Energy-Efficient Gait Adaptation via Hierarchical Reinforcement Learning for Quadrupedal Locomotion Across Diverse Terrains

Authors: Ammar Issa, Anubhav Singh, Anton Tsaritsin, Sergey Kolyubin

Organizations: Faculty of Control Systems and Robotics, ITMO University, and the Biomechatronics and Energy-Efficient Robotics Lab (BE2R), St. Petersburg, Russia

Abstract

While energy efficiency is a critical objective for legged-robot locomotion control, achieving low energy consumption while maintaining robust performance across different velocity ranges and terrain conditions remains a key challenge. This is particularly true for end-to-end RL policies, where gait generation, motion execution, and energy optimization are tightly coupled, leading to high sensitivity to reward design. In this work, we propose a hierarchical reinforcement learning (HRL) framework that separates a high-frequency policy for stable and robust joint-level motion execution from low-frequency gait adaptation that explicitly minimizes the cost of transport (CoT). The three-stage Isaac-based training procedure enables zero-shot sim-to-real transfer with improved tracking accuracy, robustness, and energy efficiency. The learned hierarchy exhibits automatic speed-dependent gait adaptation, transitioning from pacing at low speeds to trotting at higher speeds. We validate the proposed approach in simulation against representative single-policy and hierarchical locomotion baselines, demonstrating reduced CoT over a broad range of commanded velocities, while maintaining robust locomotion across flat, uneven rough, and inclined terrains. We further demonstrate its practical feasibility through zero-shot deployment on a physical Unitree AlienGo quadruped.

Figures & tables

Explore similar work

CardsList
  1. LoComposition: Terrain-Adaptive Energy-Efficient Quadruped Locomotion without Gait Priors

    Jun 14, 2026Loukas Kordos, Leonard T. Franz, Simon Rappenecker +4EmergencePrior Knowledge Integration

  2. Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion

    Jul 20, 2026Jordan Dowdy, Jean Chagas VazLegged RobotsLocomotion

  3. CoRe-MoE: Contrastive Reweighted Mixture of Experts for Multi-Terrain Humanoid Locomotion with Gait Adaptation

    Jun 3, 2026Kailun Huang, Zikang Xie, Yanzhe Xie +8Agile LocomotionGait Dynamics