Pith. sign in

Paper Citation Record · LEDGER

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving

As of 20 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2411.14716.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14716 v2

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:05:27.362145Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-17T01:59:16.251640Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T02:01:25.681466Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6aa18881-9c91-4198-9ba5-ec4a4570edb5 · outbound

This paper cites ALSO: automotive lidar self- supervision by occupancy estimation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving ALSO: automotive lidar self- supervision by occupancy estimation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.990546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.162166Z digest=sha256:cb57d9a4eddd8553333229c59459c71e6ead8af72cc1c2d3e5d7466145f017cb

Observation 44b446bc-cc5f-4bb7-9ba8-5a0dba7f7946 · outbound

This paper cites Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.979309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.166625Z digest=sha256:875465bb767f5465c6f21fc43e8b204e53f00b4135e220528418320d88c4a1e5

Observation 32523c7e-23e3-449b-9568-9bbac70847ef · outbound

This paper cites GaussianBeV: 3D Gaussian Representation meets Perception Models for BeV Segmentation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianBeV: 3D Gaussian Representation meets Perception Models for BeV Segmentation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.170569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.170569Z digest=sha256:56c953f30b18949fef0451ed87740985db759643ec039e7a625a2984a59f1e15

Observation 364d9dda-a1c3-405d-b6e4-669cea8f2498 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.967870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.175223Z digest=sha256:207c28546be79d02b6e66312d9c236b113e72f6a382d21d3c65f387ef0e244f6

Observation 8eea7ba1-caaa-41f6-9fae-5252d194ba42 · outbound

This paper cites Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.179157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.179157Z digest=sha256:f4c4a284827b53fb8d5d7b016c9e7ccbc12455843e97a59c8ef5f1c613a93590

Observation b38e2d77-f02b-4646-9620-213bc74216f0 · outbound

This paper cites Gaussianpro: 3d gaussian splatting with progressive propagation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Gaussianpro: 3d gaussian splatting with progressive propagation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.183375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.183375Z digest=sha256:c7ebce472746bfc5eff4a0e6ffffe32c80a3992784f364359e196d550d372831

Observation b5b72d35-31ca-4ea2-869d-422cddf16f66 · outbound

This paper cites MMDetection3D: Open- MMLab next-generation platform for general 3D object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving MMDetection3D: Open- MMLab next-generation platform for general 3D object detection

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.187557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.187557Z digest=sha256:5ddf1da860285da11e323e793f8f10b0b3db99c3259811bc1bce694b7be09bb4

Observation fedf35c9-25ba-4600-b167-292b5e65e2b1 · outbound

This paper cites an unresolved cited work.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T15:05:27.942940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.191103Z digest=sha256:6f85ed48a4c39f7d286bfd32ec367fb420fb23abc9f66d5f9cf5f0a014d6024e

Observation b3ff89b9-ad5a-455f-9514-d4d324f85172 · outbound

This paper cites GaussianOcc: Fully Self-supervised and Efficient 3D Occupancy Estimation with Gaussian Splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianOcc: Fully Self-supervised and Efficient 3D Occupancy Estimation with Gaussian Splatting

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.194774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.194774Z digest=sha256:781ce043d4ea0e88d8ac525d796bce2d14bc574865934a6c5383d7ca097768ab

Observation 9b0b4069-e556-435b-b94c-37a0b2dac1f7 · outbound

This paper cites Digging into self-supervised monocular depth estimation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Digging into self-supervised monocular depth estimation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.198601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.198601Z digest=sha256:594a1c17fd99319e12afc0eea432286c4e407895c5e2ae3448772ba1c875474e

Observation 916b2e7e-76e6-4e23-be54-de7b1ffacc25 · outbound

This paper cites Planning-oriented autonomous driv- ing.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Planning-oriented autonomous driv- ing

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.925559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.202410Z digest=sha256:e331965008c70cfcfe502e4222b1eee7244d44434744c2de2e9f8f68c93dc182

Observation 0eddf914-42da-4150-ab07-76c9ac3658fe · outbound

This paper cites BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVDet: High-performance Multi-camera 3D Object Detection in Bird-Eye-View

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.206211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.206211Z digest=sha256:9c60416a59888246bd18d33a2ccc9a4f41f51eacf75e553208673eedfcd7ee6b

Observation c335db95-c393-495e-8087-40c28c78c88a · outbound

This paper cites Tri-perspective view for vision- based 3d semantic occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Tri-perspective view for vision- based 3d semantic occupancy prediction

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.913850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.210238Z digest=sha256:9ac7163d4eaad990a1aa7119b5746167b14eff6c33ae6de2e160cf7286b889c3

Observation beac3a8e-e369-458b-96cf-09dd9cd43e12 · outbound

This paper cites GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.213896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.213896Z digest=sha256:88aaf9997dd3d9039efdd048c794ad2146c0c32b87d9aea30a3f3810dabc7ba3

Observation 67d713cc-7e4f-4c58-9197-877d90925caf · outbound

This paper cites Nerf-mae: Masked autoencoders for self-supervised 3d representation learning for neural radiance fields.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Nerf-mae: Masked autoencoders for self-supervised 3d representation learning for neural radiance fields

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.901390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.218037Z digest=sha256:a7b22b8431d6713edea16b8291cfdbf62d41139caf862bdb7b28f00616225d42

Observation f3c37000-0a54-4416-8f8d-99939aeb0ad4 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving 3d gaussian splatting for real-time radiance field rendering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.221978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.221978Z digest=sha256:3f1a357d1f0218396ff215bd524af01611101c5d16a8df4f9b035f85c795e309

Observation 5e45081c-65f6-4f5a-9663-f4f21306d076 · outbound

This paper cites Maeli: Masked autoencoder for large-scale lidar point clouds.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Maeli: Masked autoencoder for large-scale lidar point clouds

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.883157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.226043Z digest=sha256:cb46d09d787f0c3f10a73ba2af2380196a73996cd7766b2a807efbbde388ce73

Observation 95bb179d-5273-4705-83c1-5e3fe8d5889a · outbound

This paper cites Unifying voxel-based representation with transformer for 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unifying voxel-based representation with transformer for 3d object detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.872005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.229982Z digest=sha256:ebfdad33397b468140c772845df47541e5c2c2897ac5b80b1f6eafd5dd327db1

Observation eac0eb53-c7c9-4304-ba0b-0ac1bb817413 · outbound

This paper cites Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevdepth: Acquisition of reliable depth for multi-view 3d object detec- tion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.233686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.233686Z digest=sha256:5c289ab57cda3f653351c20979dd3999999deaf07189f0f8858cfc077e79dc4c

Observation 4e9b9075-1a1d-4b0c-a118-522fdd1133b3 · outbound

This paper cites Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Simipu: Simple 2d image and 3d point cloud un- supervised pre-training for spatial-aware visual representa- tions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.855036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.237676Z digest=sha256:5b8a1b2431f8b08be6559e8406d6a8c23a664e6f0fcbc1edc5fbacd5e730a05f

Observation 43ea415c-9bc2-49c8-878f-f7f5bfa6db9b · outbound

This paper cites Bevformer: 9 Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevformer: 9 Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.844433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.241236Z digest=sha256:3bb594524c1730a074387f845c98b812ec069ec426dd7397a70200e479dedb90

Observation 098e1ebf-13ae-4cd8-8efe-c414cfbc80be · outbound

This paper cites Fb-occ: 3d occupancy prediction based on forward-backward view transformation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fb-occ: 3d occupancy prediction based on forward-backward view transformation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.833377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.245115Z digest=sha256:89fb04b5d93ce3d5ed2f24125152883a5dd4d01dedfe83d1a0c828e009eaa08a

Observation 377dcf41-c935-4ebc-857b-33e47c637523 · outbound

This paper cites Fully sparse 3d occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fully sparse 3d occupancy prediction

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.821803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.248649Z digest=sha256:a2554f8c63731c7dadf9398e445bf7ffdcce93a0f46e35ee980170a8149953c0

Observation 86fd28be-248a-4547-a518-7b5ac43312eb · outbound

This paper cites PETR: position embedding transformation for multi-view 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving PETR: position embedding transformation for multi-view 3d object detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.809763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.252268Z digest=sha256:14ee45b286683e2ca448ffefbd99fd749e37e9e8b0095f764ade5b804fa3ca15

Observation 11d79bf9-fe4e-4747-ae37-6c204346d9d6 · outbound

This paper cites A convnet for the 2020s.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving A convnet for the 2020s

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.797683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.255728Z digest=sha256:1fb0d62d81ea234b06df8edc88f3199b70cefc8cbc463d8884e1e99b8f86486d

Observation 1968de2e-ded2-4e2e-b29e-bb6204b7e931 · outbound

This paper cites Learning ego 3d representation as ray tracing.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Learning ego 3d representation as ray tracing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.785592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.259402Z digest=sha256:594254b86b779bda1af785da7a3b9c35b35718263f1ff6f8a5f3b367642173b0

Observation 0a75d406-cd44-413e-9288-0057978490ba · outbound

This paper cites Srinivasan, Matthew Tancik, Jonathan T.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Srinivasan, Matthew Tancik, Jonathan T

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.772732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.262909Z digest=sha256:8fec853f27dbeb72a37a5788c38bde6efddc893ab9fa40db8f9216425975cdad

Observation d3440214-7ec2-433a-9352-7229afe9559a · outbound

This paper cites Occupancy-MAE: Self-supervised Pre-training Large-scale LiDAR Point Clouds with Masked Occupancy Autoencoders.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occupancy-MAE: Self-supervised Pre-training Large-scale LiDAR Point Clouds with Masked Occupancy Autoencoders

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.266517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.266517Z digest=sha256:1b871e278f5f93d6811a40db333ce974790debde3a5eb7987611ca16f4441cb3

Observation e9304f3a-0c36-415f-ac8e-3e88144a545e · outbound

This paper cites Occupancy-mae: Self-supervised pre-training large- scale lidar point clouds with masked occupancy autoen- coders.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occupancy-mae: Self-supervised pre-training large- scale lidar point clouds with masked occupancy autoen- coders

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.759678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.269898Z digest=sha256:ef0efccb45bd08c372be7658e5c680d0145d455067b87ccb6e33d2b38fe4a0e4

Observation e62b69de-5654-4eb0-a3f0-1b139c2b025b · outbound

This paper cites Segcontrast: 3d point cloud feature representation learning through self-supervised seg- ment discrimination.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Segcontrast: 3d point cloud feature representation learning through self-supervised seg- ment discrimination

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.745396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.273588Z digest=sha256:2f7f4eacde06af210bab272d268c861177eb7473e2c9b66ac4ea9bc50429c27d

Observation 2cb62c01-bad2-4606-a181-191bf402a020 · outbound

This paper cites Renderocc: Vision-centric 3d occupancy pre- diction with 2d rendering supervision.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Renderocc: Vision-centric 3d occupancy pre- diction with 2d rendering supervision

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.733778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.277305Z digest=sha256:c5816da46daa9d53ed10ba0a15e5c754e0e8821e24cf5aa779e812fa5f3d5d0d

Observation bf5ff595-6865-4b10-a764-9f8e63014d23 · outbound

This paper cites Is pseudo-lidar needed for monocular 3d object detection? In Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pages 3142–3152,.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Is pseudo-lidar needed for monocular 3d object detection? In Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pages 3142–3152,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.280812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.280812Z digest=sha256:1e312f9069f335489d6c92cc05ce36b430bb4cefdf684d7c4df12e95b63ac54a

Observation 65fbab39-a80b-4f6d-86b5-af7c3848ba3b · outbound

This paper cites BEVContrast: Self-supervision in bev space for automotive lidar point clouds.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVContrast: Self-supervision in bev space for automotive lidar point clouds

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.713543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.284207Z digest=sha256:bd22d7a68645ad8968de9410777037fec71f1f46008ac85cd2f395428b9c1c85

Observation d216e97f-d406-4719-89f6-905f7aed8f78 · outbound

This paper cites 3dppe: 3d point positional encoding for multi-camera 3d ob- ject detection transformers.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving 3dppe: 3d point positional encoding for multi-camera 3d ob- ject detection transformers

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.700033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.287586Z digest=sha256:5d2f76c2f3e478fef7e253b554540b9a8f3e44ea2d1582f03633e49c014284c1

Observation 5b95f62a-1da0-4832-b2e3-5134d95c2f1f · outbound

This paper cites Sparseocc: Re- thinking sparse latent representation for vision-based seman- tic occupancy prediction.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Sparseocc: Re- thinking sparse latent representation for vision-based seman- tic occupancy prediction

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.687493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.291176Z digest=sha256:ddfe339a37f14071d042b18bc5f3deaaffde44352f16c57bd6057d2e6c03b4d3

Observation 72f9ee61-f12f-4a46-9bf9-4ece5f3170d5 · outbound

This paper cites Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.675621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.294803Z digest=sha256:66cf449cc294654bd3bed9440198ea5e5adb5dcea804efd66bd4035be3d08d4a

Observation f5f41528-31e7-4f6e-b302-11c24e52979f · outbound

This paper cites Scene as occupancy.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Scene as occupancy

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.298136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.298136Z digest=sha256:4469a88ad88c067f79f9b040412e64fbbb0371dab61b7670b11c5bf85bb5fbbd

Observation 5198f398-2c1f-45a7-88e7-04ec9d3df602 · outbound

This paper cites Opus: occupancy prediction using a sparse set.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Opus: occupancy prediction using a sparse set

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.655341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.301497Z digest=sha256:0014aaa04840c0fe158af88a33aa4852bfd9ab12792890598940f5ea5d5920d3

Observation bd4eaec4-fa88-449a-9cb0-7cc9b50c78b0 · outbound

This paper cites Fcos3d: Fully convolutional one-stage monocular 3d object detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Fcos3d: Fully convolutional one-stage monocular 3d object detection

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.643571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.305041Z digest=sha256:373db4cd2c8c9994a29241d4fb515d04153dbfb61e95be9197f9b90eedbf4576

Observation 6f9e0367-fec3-43f0-a83f-8fb675511538 · outbound

This paper cites Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.629682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.308616Z digest=sha256:97f8402274e5fed29fadf3c55024e62332666ce6dfc524eb4e4718472f2a1311

Observation befcdced-425b-4aa2-8bbc-c27033a9dd55 · outbound

This paper cites Cross modal transformer via coordinates encoding for 3d object dectection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Cross modal transformer via coordinates encoding for 3d object dectection

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.616552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.312339Z digest=sha256:7fe9b024805863ddc000c46cf5fe71a900e4edb16d91e008580d2971c86d599a

Observation 656bbc4b-a1d4-4b23-bd45-445b060157d9 · outbound

This paper cites SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-12T15:05:27.459509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.316012Z digest=sha256:659daa9c876f71655f5b66ca97f6cc94ef349f4069fb627223e6c45eb15f918d

Observation 8d289ecb-128c-4847-b39b-9a2608a3034e · outbound

This paper cites Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.319583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.319583Z digest=sha256:426644753096c2c50bd9fb85afa0c3a17b63a1f03d268ca74ed0455612497a6a

Observation 2e8f9ebd-155d-4b75-bfd3-bd65ba0b3416 · outbound

This paper cites Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.323607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.323607Z digest=sha256:ea4220a04a42730f8645a49796f8ad7dc306c48430e3817bdf4d72ab5316de2e

Observation 4707a4ac-6400-442b-990b-eb6ce7df8164 · outbound

This paper cites Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective supervision.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective supervision

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.603727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.328573Z digest=sha256:c98a4e67eb55422f3aa2c9466282067bbe7b7aafb3ab4872945063a58987ac4b

Observation c9c3ce52-b2ec-4e16-8d1d-0c078af174ff · outbound

This paper cites Unipad: A universal pre-training paradigm for autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Unipad: A universal pre-training paradigm for autonomous driving

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.589686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.332025Z digest=sha256:886597c1b49c7628bc2861e141904f5c0aaf6bc9bf3ae8f42f8919d7319ca1c9

Observation 7fc07850-f963-4f17-8002-fa91e2da01d7 · outbound

This paper cites Visual point cloud forecasting enables scalable autonomous driving.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Visual point cloud forecasting enables scalable autonomous driving

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.577009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.335585Z digest=sha256:9add17c111a71a79bd31f864f3df6b896c74e8bdf52230be84d765e13f414a34

Observation 94a9b3b0-098f-4769-9ab1-63ec187a6663 · outbound

This paper cites Ad-pt: Autonomous driving pre-training with large-scale point cloud dataset.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Ad-pt: Autonomous driving pre-training with large-scale point cloud dataset

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.565893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.338784Z digest=sha256:3d0b9653c18cae9717db0021618e86f287fb0536f36f66e7d7aaab64ea358441

Observation 5a445193-deb4-44c1-8a0f-1fded900b4ce · outbound

This paper cites BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.342139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.342139Z digest=sha256:ee5ca1eac3c51179d98e2c08f48466c4fe3ababfd58e8dba2e6062bb22822cd3

Observation 1c66db66-0e73-4b5f-899c-dee36e580ca5 · outbound

This paper cites Radocc: Learning cross-modality occupancy knowledge through ren- dering assisted distillation.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Radocc: Learning cross-modality occupancy knowledge through ren- dering assisted distillation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.554208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.346075Z digest=sha256:15df178dd2450665a85714b7ada920f67d5599f6391e7e2b33038c2c600043ea

Observation 7a88f6a2-edf0-49e9-83e9-c4a5585f7d03 · outbound

This paper cites Hugs: Holistic urban 3d scene understanding via gaus- sian splatting.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Hugs: Holistic urban 3d scene understanding via gaus- sian splatting

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:05:27.541582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.350412Z digest=sha256:5e3bcc0ccd2b7e9ee8057dc88c1b7d1de71ba38137520129ee876b4ef3170dca

Observation 02d46c52-78ae-4cf2-8904-252f418a856d · outbound

This paper cites Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.354297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.354297Z digest=sha256:a91a1663ce7210f37606dd9e908b353276453b24df2d2b4d64b5e8e55c188a68

Observation d8316332-66dd-46a4-8b7b-6ca159295ba6 · outbound

This paper cites Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving Class-balanced Grouping and Sampling for Point Cloud 3D Object Detection

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T15:05:27.358066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:05:27.358066Z digest=sha256:73ac4f9730b9bbe486eafd4bfe1b876938bb8180bfb75ae302ace8abe0599ff6

Observation 92d8d6a5-fac6-4d1e-96b3-b4d0aea5bc2c · outbound

This paper cites MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning.

VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-12T15:05:27.399830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T15:05:27.362145Z digest=sha256:31fa9fe804b95b4129d2738febb6100b91c6a93d2681f5928712c505a53b46d6

Pith citing papers

Observation 42612af9-ea0c-4ba2-b543-5b219ee4f4d0 · inbound

Flux4D: Flow-based Unsupervised 4D Reconstruction cites this paper.

Flux4D: Flow-based Unsupervised 4D Reconstruction VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:01:25.683602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T01:59:16.251640Z digest=sha256:8a59d21fbe2dbedde9683862f4a32640a1de24577e81ead3f44344d567a4cfb2