Learning and Deploying Robust Locomotion Policies with Minimal Dynamics Randomization @OxfordDynamicRobotSystemsGroup
Learning and Deploying Robust Locomotion Policies with Minimal Dynamics Randomization  @OxfordDynamicRobotSystemsGroup
Uploaded September 2022 | Updated September 2026, 3 weeks ago
Training deep reinforcement learning (DRL) locomotion policies often requires massive amounts of data to converge to the desired behavior. In this regard, simulators provide a cheap and abundant source. For successful sim-to-real transfer, exhaustively engineered approaches such as system identification, dynamics randomization, and domain adaptation are generally employed. As an alternative, we investigate a simple strategy of random force injection (RFI) to perturb system dynamics during training. We show that the application of random forces enables us to emulate dynamics randomization. This allows us to obtain locomotion policies that are robust to variations in system dynamics. We further extend RFI, referred to as extended random force injection (ERFI), by introducing an episodic actuation offset. We demonstrate that ERFI provides additional robustness for variations in system mass offering on average a 61% improved performance over RFI. We also show that ERFI is sufficient to perform a successful sim-to-real transfer on two different quadrupedal platforms, ANYmal C and Unitree A1, even for perceptive locomotion over uneven terrain in outdoor environments.

Authors: Luigi Campanaro, Siddhant Gangapurwala, Wolfgang Merkt, Ioannis Havoutis

Pre-print: arxiv.org/abs/2209.12878
Website: sites.google.com/view/erfi-icra
Learning and Deploying Robust Locomotion Policies with Minimal Dynamics RandomizationRA-L/ICRA 2021 - Unified Landmark Tracking for Odometry [Finalist ICRA Best Student Paper]Ensuring Safe Visual Teach and Repeat Legged Navigation using Local World Representations[Presentation] Elastic and Efficient LiDAR Reconstruction for Large-Scale Exploration TasksPlane Seg on ANYmal (Simulation)RA-L/IROS 2019 - Robust Legged Robot State Estimation Using Factor Graph OptimizationLearning Low-Frequency Motion Control for Robust and Dynamic Robot LocomotionSapling-NeRF: Geo-Localised Sapling Reconstruction in Forests for Ecological MonitoringSafe Visual Navigation in an Underground MineMapping Oxfords Sheldonian TheatreDeep IMU Bias Inference for Robust Visual-Inertial Odometry With Factor Graphs3D Scanning of the Chernobyl Nuclear Power Plant
Oxford Dynamic Robot Systems Group |

Learning and Deploying Robust Locomotion Policies with Minimal Dynamics Randomization

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER