Pith. sign in

Deep Patch Visual SLAM

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Recent work in visual SLAM has shown the effectiveness of using deep network backbones. Despite excellent accuracy, however, such approaches are often expensive to run or do not generalize well zero-shot. Their runtime can also fluctuate wildly while their frontend and backend fight for access to GPU resources. To address these problems, we introduce Deep Patch Visual (DPV) SLAM, a method for monocular visual SLAM on a single GPU. DPV-SLAM maintains a high minimum framerate and small memory overhead (5-7G) compared to existing deep SLAM systems. On real-world datasets, DPV-SLAM runs at 1x-4x real-time framerates. We achieve comparable accuracy to DROID-SLAM on EuRoC and TartanAir while running 2.5x faster using a fraction of the memory. DPV-SLAM is an extension to the DPVO visual odometry system; its code can be found in the same repository: https://github.com/princeton-vl/DPVO

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Princeton365: A Diverse Dataset with Accurate Camera Pose

cs.CV · 2025-06-10 · conditional · novelty 7.0

Princeton365 is a 365-video SLAM/NVS benchmark with board-calibrated millimeter-accurate 6-DoF poses, a new scale-aware optical-flow error metric, and an NVS benchmark of fully non-Lambertian 360-degree scans.

citing papers explorer

Showing 1 of 1 citing paper.

  • Princeton365: A Diverse Dataset with Accurate Camera Pose cs.CV · 2025-06-10 · conditional · none · ref 22 · internal anchor

    Princeton365 is a 365-video SLAM/NVS benchmark with board-calibrated millimeter-accurate 6-DoF poses, a new scale-aware optical-flow error metric, and an NVS benchmark of fully non-Lambertian 360-degree scans.