Pith. sign in

REVIEW 1 cited by

Self-Improving SLAM in Dynamic Environments: Learning When to Mask

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.08350 v3 pith:V6SWZLZN submitted 2022-10-15 cs.CV cs.AI

classification cs.CVcs.AI
keywords objectsslamdynamicmaskmethodwhenmaskingconsinv
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Visual SLAM - Simultaneous Localization and Mapping - in dynamic environments typically relies on identifying and masking image features on moving objects to prevent them from negatively affecting performance. Current approaches are suboptimal: they either fail to mask objects when needed or, on the contrary, mask objects needlessly. Thus, we propose a novel SLAM that learns when masking objects improves its performance in dynamic scenarios. Given a method to segment objects and a SLAM, we give the latter the ability of Temporal Masking, i.e., to infer when certain classes of objects should be masked to maximize any given SLAM metric. We do not make any priors on motion: our method learns to mask moving objects by itself. To prevent high annotations costs, we created an automatic annotation method for self-supervised training. We constructed a new dataset, named ConsInv, which includes challenging real-world dynamic sequences respectively indoors and outdoors. Our method reaches the state of the art on the TUM RGB-D dataset and outperforms it on KITTI and ConsInv datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. InCrowd-VI: A Realistic Visual-Inertial Dataset for Evaluating SLAM in Indoor Pedestrian-Rich Spaces for Human Navigation

    cs.RO 2024-11 conditional novelty 7.0 of 10

    InCrowd-VI provides a realistic visual-inertial benchmark with 58 head-worn sequences in crowded indoor spaces, where state-of-the-art SLAM systems frequently fail to meet accuracy and real-time requirements.

Pith tools