English

RuN: Residual Policy for Natural Humanoid Locomotion

Robotics 2025-09-26 v1

Abstract

Enabling humanoid robots to achieve natural and dynamic locomotion across a wide range of speeds, including smooth transitions from walking to running, presents a significant challenge. Existing deep reinforcement learning methods typically require the policy to directly track a reference motion, forcing a single policy to simultaneously learn motion imitation, velocity tracking, and stability maintenance. To address this, we introduce RuN, a novel decoupled residual learning framework. RuN decomposes the control task by pairing a pre-trained Conditional Motion Generator, which provides a kinematically natural motion prior, with a reinforcement learning policy that learns a lightweight residual correction to handle dynamical interactions. Experiments in simulation and reality on the Unitree G1 humanoid robot demonstrate that RuN achieves stable, natural gaits and smooth walk-run transitions across a broad velocity range (0-2.5 m/s), outperforming state-of-the-art methods in both training efficiency and final performance.

Keywords

Cite

@article{arxiv.2509.20696,
  title  = {RuN: Residual Policy for Natural Humanoid Locomotion},
  author = {Qingpeng Li and Chengrui Zhu and Yanming Wu and Xin Yuan and Zhen Zhang and Jian Yang and Yong Liu},
  journal= {arXiv preprint arXiv:2509.20696},
  year   = {2025}
}