Pith. sign in

REVIEW 1 cited by

RLOC: Terrain-Aware Legged Locomotion using Reinforcement Learning and Optimal Control

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.03094 v3 pith:A42DEBNY submitted 2020-12-05 cs.RO cs.LG

classification cs.ROcs.LG
keywords controllocomotionanymalcomplexevaluatefootstepgeneratedlearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present a unified model-based and data-driven approach for quadrupedal planning and control to achieve dynamic locomotion over uneven terrain. We utilize on-board proprioceptive and exteroceptive feedback to map sensory information and desired base velocity commands into footstep plans using a reinforcement learning (RL) policy. This RL policy is trained in simulation over a wide range of procedurally generated terrains. When ran online, the system tracks the generated footstep plans using a model-based motion controller. We evaluate the robustness of our method over a wide variety of complex terrains. It exhibits behaviors which prioritize stability over aggressive locomotion. Additionally, we introduce two ancillary RL policies for corrective whole-body motion tracking and recovery control. These policies account for changes in physical parameters and external perturbations. We train and evaluate our framework on a complex quadrupedal system, ANYmal version B, and demonstrate transferability to a larger and heavier robot, ANYmal C, without requiring retraining.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Reference Free Platform Adaptive Locomotion for Quadrupedal Robots using a Dynamics Conditioned Policy

    cs.RO 2025-05 conditional novelty 5.0 of 10

    A single dynamics-conditioned RL policy transfers zero-shot across quadrupeds from 12 kg to 50 kg, and diverse reference robots during training clearly improve tracking.

Pith tools