REVIEW 6 cited by
Rapid Locomotion via Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Agile maneuvers such as sprinting and high-speed turning in the wild are challenging for legged robots. We present an end-to-end learned controller that achieves record agility for the MIT Mini Cheetah, sustaining speeds up to 3.9 m/s. This system runs and turns fast on natural terrains like grass, ice, and gravel and responds robustly to disturbances. Our controller is a neural network trained in simulation via reinforcement learning and transferred to the real world. The two key components are (i) an adaptive curriculum on velocity commands and (ii) an online system identification strategy for sim-to-real transfer leveraged from prior work. Videos of the robot's behaviors are available at: https://agility.csail.mit.edu/
Forward citations
Cited by 6 Pith papers
-
Hip Energized Monopedal Hopping
A hip-actuated monoped can stabilize pitch and energize its hop with the same torque, and its steady-state gait has closed-form fixed points and eigenvalues from hybrid averaging, validated on the Penn Jerboa robot.
-
GenTrack: Physical Alignment for Robot-Native Motion Generation and Zero-Shot Humanoid Tracking
Online co-training of a text-to-motion generator and a humanoid tracker on simulated G1 improves generator executability and zero-shot tracker coverage beyond static replay or one-way filtering.
-
Q-learning-based Model-free Safety Filter
A Q-learning safety filter with a time-dependent reward blocks unsafe actions from arbitrary task policies, but its theoretical guarantee is not valid as written.
-
ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills
ASAP trains a residual action model on real-world rollouts and fine-tunes simulation policies through it, reducing humanoid whole-body motion tracking error in sim-to-real transfer.
-
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control
A decoupled humanoid controller combines IK-based arm control with an RL locomotion policy conditioned on a CVAE motion prior, improving manipulation precision while maintaining walking stability.
-
GainAdaptor: Learning Quadrupedal Locomotion with Dual Actors for Adaptable and Energy-Efficient Walking on Various Terrains
GainAdaptor learns to adjust both joint positions and PD gains with two neural network actors, cutting power use on a Unitree Go1 across varied terrains.
Discussion (0). Continue with ORCID to comment.