Pith. sign in

Paper Citation Record · LEDGER

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains

As of 8 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2506.05250.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05250 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:26:12.764423Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9de558b-9d45-4209-b08d-0eef1e6588aa · outbound

This paper cites Fine-grained cross- view geo-localization using a correlation-aware homography estimator.Advances in Neural Information Processing Systems, 36:5301–5319, 2023.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Fine-grained cross- view geo-localization using a correlation-aware homography estimator.Advances in Neural Information Processing Systems, 36:5301–5319, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:11.390530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:11.390530Z digest=sha256:033696d91f22be719f18a664f07c9b86809347ee59497a2c7a33a9faa26c2b1f

Observation f7b5ad67-3b92-4ae6-8b73-7d39ca366ae5 · outbound

This paper cites Boosting 3-dof ground- to-satellite camera localization accuracy via geometry-guided cross-view transformer.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Boosting 3-dof ground- to-satellite camera localization accuracy via geometry-guided cross-view transformer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.736711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.450142Z digest=sha256:49d79c326bbd556c4fbf0f784b0874650ac50eb4c99662f0c70df2c5ed342c34

Observation 35630bdb-f25a-4ca1-8e5b-2b64b192a407 · outbound

This paper cites Bevrender: Vision-based cross-view vehicle registration in off-road gnss-denied environment.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Bevrender: Vision-based cross-view vehicle registration in off-road gnss-denied environment

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.508209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.520044Z digest=sha256:7b0cd1337f222544e90579cf1deaf127b795b3474a7de74e054da3a18193d403

Observation b57840c7-68c2-427a-8bd9-8b0545006df2 · outbound

This paper cites Bevloc: Cross-view localization and matching via birds-eye-view synthesis.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Bevloc: Cross-view localization and matching via birds-eye-view synthesis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.414082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.562368Z digest=sha256:f325076072355391a4b79aed8ed294c72d01e4a4f36d41fb357ef34669f1224e

Observation 1612b692-ba85-43ce-acb5-c5971b69d842 · outbound

This paper cites Orienternet: Visual localization in 2d public maps with neural matching.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Orienternet: Visual localization in 2d public maps with neural matching

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.259633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.628834Z digest=sha256:2f37ad92d5ab08f4003ba64c55770a568e435391266974df0a4a35ea9e81bb1c

Observation 96b9dacc-2086-476b-96f8-c8304af5d26b · outbound

This paper cites Coming down to earth: Satellite-to-street view synthesis for geo-localization.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Coming down to earth: Satellite-to-street view synthesis for geo-localization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.145404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.705869Z digest=sha256:4df37bf9cc9e382d9635e2226201dfff42462c42909a608f66deaf9c87280eff

Observation b0bf9bb7-dc72-4fae-8fd1-e989803b313b · outbound

This paper cites Wide-area geolocalization with a limited field of view camera.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Wide-area geolocalization with a limited field of view camera

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:14.035366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.782657Z digest=sha256:f921959bb3aea9549ac9d3d36c6cdd6274c8fd93e65011c8bf8e30779fa3775d

Observation 71fe8a78-fd6a-4963-92eb-300f683fc480 · outbound

This paper cites Gama: Cross-view video geo-localization.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Gama: Cross-view video geo-localization

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.922093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.852600Z digest=sha256:4eec65d9c576e866b84f3bd209820550dcf13117e0d2e031ea4fba28389571df

Observation 3e52074c-0d58-49be-8433-3994239a6696 · outbound

This paper cites Garet: Cross-view video geolocalization with adapters and auto-regressive transformers.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Garet: Cross-view video geolocalization with adapters and auto-regressive transformers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.788999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.902471Z digest=sha256:28bf4cb7a16627ad831b1fbe8a09abf44e6d2830a37ed601007e69210e0a657f

Observation 3c69c423-b99d-4d53-982d-b67c22de96e4 · outbound

This paper cites A cross-view geo-localization algorithm using uav image and satellite image.Sensors, 24(12):3719, 2024.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains A cross-view geo-localization algorithm using uav image and satellite image.Sensors, 24(12):3719, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.675272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:11.966271Z digest=sha256:0de589f31cc1fb9298561b31e628edab636b236a6a144d9d35e074f931459dd5

Observation 9ae02d6b-2830-4434-ade0-16ab1a9c6f5a · outbound

This paper cites Accurate 3-dof camera geo-localization via ground-to-satellite image matching.IEEE transactions on pattern analysis and machine intelligence, 45(3):2682–2697, 2022.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Accurate 3-dof camera geo-localization via ground-to-satellite image matching.IEEE transactions on pattern analysis and machine intelligence, 45(3):2682–2697, 2022

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:12.023572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:12.023572Z digest=sha256:c9fe7814096bc4a2fa383c5eb32d575a335c555ac15285a917eca1a585666079

Observation 49c023f6-f82e-4cfb-88e8-b59c655c208b · outbound

This paper cites Tartandrive 2.0: More modalities and better infrastructure to further self-supervised learning research in off-road driving tasks.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Tartandrive 2.0: More modalities and better infrastructure to further self-supervised learning research in off-road driving tasks

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.564656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.074635Z digest=sha256:0f2c2224dcb0a2b085b88f659c80bf549183987c426127ba6fb12471f42e2bbe

Observation 667732df-5ef2-4e1e-b0b2-3a8613339cac · outbound

This paper cites MIT Press, 2005.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains MIT Press, 2005

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.450044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.149295Z digest=sha256:e6266452fd10739ec34a939a290402a3b56660e98b8729810529d5f00536045d

Observation 03a6374e-73ef-4a97-94c0-fc944f916dbe · outbound

This paper cites Patch-netvlad: Multi-scale fusion of locally-global descriptors for place recognition.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Patch-netvlad: Multi-scale fusion of locally-global descriptors for place recognition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:12.190607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:12.190607Z digest=sha256:ddf2a43e21b8f562926798afaa5f63c412d4f113b156d95c0bf7f0b37a026b1a

Observation f61dd4e3-3a09-49e0-aea8-dfe31569eb87 · outbound

This paper cites Boosting contrastive self-supervised learning with false negative cancellation.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Boosting contrastive self-supervised learning with false negative cancellation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.333398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.280631Z digest=sha256:d7d4debce031fff16286b758ce0a24ee987eca002a120b5cdc7e86279aab90df

Observation 7fe55479-54b6-476b-9ebe-3d082d0cb3db · outbound

This paper cites Spatiotemporal contrastive video representation learning.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Spatiotemporal contrastive video representation learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.250914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.354674Z digest=sha256:6b11d691d3a0473a9f7a9730e48dbcdef3053e51636a66d59ca8bb61a1a29151

Observation 9d0892d0-a9fe-44ca-8bb0-535f8f388a7b · outbound

This paper cites Kld-sampling: Adaptive particle filters.Advances in neural information processing systems, 14, 2001.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Kld-sampling: Adaptive particle filters.Advances in neural information processing systems, 14, 2001

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.167846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.427815Z digest=sha256:f5d4fa8d50ed2abcb54434a7a7d3884e3b7ae5a05b177aca1edfc63a8b118721

Observation e572c83b-a572-4bfd-93cb-bbd01d7431c5 · outbound

This paper cites Google maps static api, 2024.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Google maps static api, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.086653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.461911Z digest=sha256:eade978d7fcdf499c0aec9947a0ce7b6c028022fef3e3a8f28ff26a945b8e250

Observation 56690b08-6010-40cd-becf-0b9584b1ee66 · outbound

This paper cites Improved deep metric learning with multi-class n-pair loss objective.Advances in neural information processing systems, 29, 2016.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Improved deep metric learning with multi-class n-pair loss objective.Advances in neural information processing systems, 29, 2016

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:13.019255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.522759Z digest=sha256:27895ba83857263874866a9def111e0d7deb2373473809afc0a183e841a6d794

Observation 42b8c5a3-86d1-4711-aeae-a138644696e1 · outbound

This paper cites Tartanvo: A generalizable learning-based vo.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Tartanvo: A generalizable learning-based vo

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:12.593547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:12.593547Z digest=sha256:7d71138cd59357a46c42b590b37630c696640aead3a0497d2532a0a6765bc713

Observation 57825ca0-3074-45a5-bed3-e3bdc45a730a · outbound

This paper cites Super odometry: Imu-centric lidar-visual-inertial estimator for challenging environments.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains Super odometry: Imu-centric lidar-visual-inertial estimator for challenging environments

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:26:12.925905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:26:12.657518Z digest=sha256:4970fd3694415834d2c54d0afc39efb67529c2a6ee2f6249e46d6c99cf42f03e

Observation 604f17fe-3431-4d05-94dd-afe6b4e4d60c · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:12.720541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:12.720541Z digest=sha256:84ca2020294622a5d773772c1d75b8dc23459c283fd917bcb0b085d1b521eefe

Observation eefb555b-a38a-4792-91c4-8393e91b172d · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Spatiotemporal Contrastive Learning for Cross-View Video Localization in Unstructured Off-road Terrains DINOv2: Learning Robust Visual Features without Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:26:12.764423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:26:12.764423Z digest=sha256:be42d1b80ff3c2ff629f32599508e3475e402ee1d7da343c4a568e2edda67946

Pith citing papers

No inbound Pith citation observations are available.