English

Learning Velocity-based Humanoid Locomotion: Massively Parallel Learning with Brax and MJX

Robotics 2024-07-09 v1

Abstract

Humanoid locomotion is a key skill to bring humanoids out of the lab and into the real-world. Many motion generation methods for locomotion have been proposed including reinforcement learning (RL). RL locomotion policies offer great versatility and generalizability along with the ability to experience new knowledge to improve over time. This work presents a velocity-based RL locomotion policy for the REEM-C robot. The policy uses a periodic reward formulation and is implemented in Brax/MJX for fast training. Simulation results for the policy are demonstrated with future experimental results in progress.

Keywords

Cite

@article{arxiv.2407.05148,
  title  = {Learning Velocity-based Humanoid Locomotion: Massively Parallel Learning with Brax and MJX},
  author = {William Thibault and William Melek and Katja Mombaur},
  journal= {arXiv preprint arXiv:2407.05148},
  year   = {2024}
}
R2 v1 2026-06-28T17:31:29.035Z