cs.ROOct 1, 2026

MASkillBlender: Decentralized Whole-Body Coordination for Multi-Humanoid Loco-Manipulation via Skill Blending

Authors: Yifan Hu, Luhang Hong, Mingkang Long, Danning Wang, Chengfeng Jia, Rong Su, Junjie Fu, Guanghui Wen

Organizations: Nanyang Technological University · Southeast University · Purple Mountain Laboratories

Abstract

Coordinated multi-humanoid loco-manipulation is promising yet challenging due to high-dimensional whole-body control, decentralized decision making, and scalability. While recent reinforcement learning methods have improved single-humanoid whole-body control, extending them to the multi-humanoid setting remains nontrivial and often requires substantial reward engineering or task-specific design. We propose MASkillBlender, a general multi-agent reinforcement learning framework to achieve decentralized multi-humanoid whole-body coordination. By learning a shared decentralized high-level policy over reusable pre-trained single-humanoid skills, MASkillBlender enables coordinated behaviors using only task-level rewards, without requiring task-specific motion references. To improve learning efficiency, we further introduce a permutation-based data augmentation strategy for homogeneous multi-humanoid systems, and theoretically show that the permuted samples preserve the policy-gradient direction of the original samples under the homogeneous Markov game formulation. We evaluate MASkillBlender on multiple multi-humanoid coordination tasks across two humanoid embodiments. Simulation results demonstrate that the proposed framework consistently achieves strong task performance and enables coordinated behaviors across different tasks and humanoid embodiments.

Figures & tables

Appendix figures & tables11 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. SkillX: Unified Multi-Skill Policy Learning for Humanoid Soccer

    Date pendingZhangchen Ye, Enxuan Ruan, Yifei Bao +10BadmintonHuman-To-Robot Transfer

  2. Distilling Collaborative Dynamics into Latent Space for Implicit Coordination in Decentralized Multi-Agent Manipulation

    Jun 22, 2026Chanyoung Park, Minsung Yoon, Andrew Jeong +1Multi-Agent CoordinationGlobal Coordination

  3. Learning Multi-Humanoid Pickup and Transport via Decentralized Object-Centric Control

    Sep 15, 2026Bikram Pandit, Mohitvishnu S. Gadde, Aayam Kumar Shrestha +1Bimanual ManipulationReal-World Manipulation Tasks