Pith. sign in

REVIEW 2 cited by

Learning to Jump from Pixels

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.15344 v1 pith:DV7X7T5W submitted 2021-10-28 cs.RO cs.AI

classification cs.ROcs.AI
keywords agilebehaviorschallengescontroldiscontinuouslearninglocomotionmethod
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Today's robotic quadruped systems can robustly walk over a diverse range of rough but continuous terrains, where the terrain elevation varies gradually. Locomotion on discontinuous terrains, such as those with gaps or obstacles, presents a complementary set of challenges. In discontinuous settings, it becomes necessary to plan ahead using visual inputs and to execute agile behaviors beyond robust walking, such as jumps. Such dynamic motion results in significant motion of onboard sensors, which introduces a new set of challenges for real-time visual processing. The requirement for agility and terrain awareness in this setting reinforces the need for robust control. We present Depth-based Impulse Control (DIC), a method for synthesizing highly agile visually-guided locomotion behaviors. DIC affords the flexibility of model-free learning but regularizes behavior through explicit model-based optimization of ground reaction forces. We evaluate the proposed method both in simulation and in the real world.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?

    cs.CV 2024-12 conditional novelty 6.0 of 10

    State-to-Visual DAgger outperforms visual RL on hard manipulation tasks and is more stable and faster in wall-clock time, but offers little sample-efficiency benefit on easy tasks.

  2. Reinforcement Learning from Wild Animal Videos

    cs.RO 2024-12 conditional novelty 6.0 of 10

    A quadruped robot acquires walking, jumping, running-like, and standing skills using only the output of a video classifier trained on wild-animal videos as its reinforcement learning reward.

Pith tools