REVIEW 2 cited by
Learning Risk-Aware Costmaps via Inverse Reinforcement Learning for Off-Road Navigation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The process of designing costmaps for off-road driving tasks is often a challenging and engineering-intensive task. Recent work in costmap design for off-road driving focuses on training deep neural networks to predict costmaps from sensory observations using corpora of expert driving data. However, such approaches are generally subject to over-confident mispredictions and are rarely evaluated in-the-loop on physical hardware. We present an inverse reinforcement learning-based method of efficiently training deep cost functions that are uncertainty-aware. We do so by leveraging recent advances in highly parallel model-predictive control and robotic risk estimation. In addition to demonstrating improvement at reproducing expert trajectories, we also evaluate the efficacy of these methods in challenging off-road navigation scenarios. We observe that our method significantly outperforms a geometric baseline, resulting in 44% improvement in expert path reconstruction and 57% fewer interventions in practice. We also observe that varying the risk tolerance of the vehicle results in qualitatively different navigation behaviors, especially with respect to higher-risk scenarios such as slopes and tall grass.
Forward citations
Cited by 2 Pith papers
-
Discriminative Barrier Functions for Safe Adversarial Imitation Learning from Observation
Constraining the adversarial imitation learning discriminator to discrete-time control barrier functions recovers safety barriers from unlabeled observations and reduces collisions in navigation.
-
Implicit Dual-Control for Visibility-Aware Navigation in Unstructured Environments
VA-MPPI is a model predictive path integral controller that uses predicted visibility to update terrain uncertainty inside each rollout, showing in simulation fewer collisions in occluded environments than a determini...
Discussion (0). Sign in to comment.