REVIEW 5 cited by
Where is your place, Visual Place Recognition?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Where is your place, Visual Place Recognition?
read the original abstract
Visual Place Recognition (VPR) is often characterized as being able to recognize the same place despite significant changes in appearance and viewpoint. VPR is a key component of Spatial Artificial Intelligence, enabling robotic platforms and intelligent augmentation platforms such as augmented reality devices to perceive and understand the physical world. In this paper, we observe that there are three "drivers" that impose requirements on spatially intelligent agents and thus VPR systems: 1) the particular agent including its sensors and computational resources, 2) the operating environment of this agent, and 3) the specific task that the artificial agent carries out. In this paper, we characterize and survey key works in the VPR area considering those drivers, including their place representation and place matching choices. We also provide a new definition of VPR based on the visual overlap -- akin to spatial view cells in the brain -- that enables us to find similarities and differences to other research areas in the robotics and computer vision fields. We identify numerous open challenges and suggest areas that require more in-depth attention in future works.
Forward citations
Cited by 5 Pith papers
-
Breaking D\'ej\`a Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning
VLM-based post-retrieval auditing of visual place recognition raises recall@1 by 13.6% on average while cutting false accepts to 12% and holding precision above 95%.
-
Breaking D\'ej\`a Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning
A zero-shot VLM auditor that accepts or rejects top-1 VPR matches is reported to improve recall@1 by 13.6% on average, but the gain appears to depend on a nonstandard filtered-recall definition.
-
Defending from GeoLocalization through Adversarial Road Trips
RoadTrip Attack uses beam search over adaptive geographic intermediate targets to produce stronger, more transferable, lower-visibility adversarial examples against retrieval-based image geolocalizers than PGD, FGSM, ...
-
Long-Term Visual Localization in Dynamic Benthic Environments: A Dataset, Footprint-Based Ground Truth, and Visual Place Recognition Benchmark
A benchmark shows state-of-the-art visual place recognition performs poorly on a new multi-site, multi-year benthic AUV dataset, and that distance-based ground truth inflates recall.
-
UniPR-3D: Towards Universal Visual Place Recognition with Visual Geometry Grounded Transformer
A VGGT-based descriptor merging 2D and 3D transformer tokens sets new state-of-the-art recall on single- and multi-frame visual place recognition benchmarks.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.