REVIEW 14 cited by
Learning Humanoid Locomotion over Challenging Terrain
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Humanoid robots can, in principle, use their legs to go almost anywhere. Developing controllers capable of traversing diverse terrains, however, remains a considerable challenge. Classical controllers are hard to generalize broadly while the learning-based methods have primarily focused on gentle terrains. Here, we present a learning-based approach for blind humanoid locomotion capable of traversing challenging natural and man-made terrain. Our method uses a transformer model to predict the next action based on the history of proprioceptive observations and actions. The model is first pre-trained on a dataset of flat-ground trajectories with sequence modeling, and then fine-tuned on uneven terrain using reinforcement learning. We evaluate our model on a real humanoid robot across a variety of terrains, including rough, deformable, and sloped surfaces. The model demonstrates robust performance, in-context adaptation, and emergent terrain representations. In real-world case studies, our humanoid robot successfully traversed over 4 miles of hiking trails in Berkeley and climbed some of the steepest streets in San Francisco.
Forward citations
Cited by 14 Pith papers
-
In vivo feasibility study of humanoid robots in surgery
Teleoperated Unitree G1 humanoids using manual wristed instruments completed two in-vivo porcine cholecystectomies and dry-lab tasks with performance between manual laparoscopy and commercial surgical robots.
-
Visual Imitation Enables Contextual Humanoid Control
A single policy trained from 123 monocular videos, fine-tuned in simulation, and distilled to heightmap plus root-direction inputs lets a Unitree G1 climb stairs and sit and stand on real furniture.
-
MuJoCo Playground
An open-source, MJX-based robot learning framework with integrated batch rendering that provides fast training and demonstrates sim-to-real transfer on six robot platforms.
-
Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains
A proprioceptive humanoid policy trained with slope-adaptive ZMP regularization plus biomechanical reward gating traverses outdoor grass slopes to 32.1° without online exteroception.
-
PHUMA: Physically Reliable Humanoid Locomotion Dataset
PHUMA is a curated 73-hour humanoid locomotion corpus whose physical-reliability metrics are partly defined by the same losses used to optimize it, and whose imitation success claims are confounded by in-distribution ...
-
KLEIYN : A Quadruped Robot with an Active Waist for Both Locomotion and Wall Climbing
KLEIYN, a quadruped with an active waist joint, climbs chimneys 800-1000 mm wide at roughly 150 mm/s using a reinforcement learning curriculum that eases the wall-floor transition from curved to vertical.
-
GMT: General Motion Tracking for Humanoid Whole-Body Control
GMT trains a single unified humanoid policy using adaptive sampling and mixture-of-experts, achieving lower tracking errors than a re-implemented ExBody2 across diverse whole-body motions.
-
TWIST: Teleoperated Whole-Body Imitation System
Human MoCap drives a Unitree G1 humanoid in real time through a single teacher-student RL+BC controller that transfers zero-shot from simulation.
-
BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds
A two-stage RL framework with a polygonal-foot foothold reward and double critic enables a Unitree G1 humanoid to traverse sparse footholds in simulation and the real world.
-
Learning from Massive Human Videos for Universal Humanoid Pose Control
Humanoid-X contributes 163,800 text-annotated motion clips retargeted from human videos into humanoid robot poses, and UH-1 is an autoregressive transformer that maps text instructions to humanoid actions.
-
Semantic Audio-driven Understanding for Dynamic Humanoid Whole Body Control
A multi-modal audio router maps streaming music and speech to imitation-learned whole-body policies for a Unitree G1 humanoid, achieving 84.8% chunk-level retrieval accuracy in simulation.
-
Whole-Body Conditioned Egocentric Video Prediction
An autoregressive conditional diffusion transformer predicts future egocentric video from whole-body 3D pose sequences, trained on Nymeria, with atomic action and long-horizon evaluations.
-
HiLo: Learning Whole-Body Human-like Locomotion with Motion Tracking Controller
A humanoid locomotion controller that combines open-loop reference tracking with a residual RL policy transfers to the real robot and changes gait style by reweighting the reference motion.
-
Robust RL Control for Bipedal Locomotion with Closed Kinematic Chains
A reinforcement-learning gait controller that explicitly models closed kinematic chains outperforms one trained on a simplified serial model, both in simulation and on the physical TopA robot.
Discussion (0). Continue with ORCID to comment.