Pith. sign in

Paper Citation Record · LEDGER

BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2203.17270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.17270 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:47:13.263203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:08:22.583557Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e62581fb-6d06-48b9-b7a1-b77bca286d76 · inbound

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning cites this paper.

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:18:49.535690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T03:16:03.058129Z digest=sha256:80200bdde6a912df28c119b4349d8854013303b277732503c4d5a2fc1415aee2

Observation b43c0cbb-c37b-456c-8b1c-44108f63848e · inbound

Enhancing End-to-End Autonomous Driving with Latent World Model cites this paper.

Enhancing End-to-End Autonomous Driving with Latent World Model BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:38:51.943862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T07:38:51.885383Z digest=sha256:a5033818ac36932316870e5c3aef18c10c2b556b9f187fb806eaaca27fcc39d2

Observation 0cd24a7c-0bac-45db-8965-9174ce6315cf · inbound

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving cites this paper.

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:25:51.484526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T21:25:20.390835Z digest=sha256:e3e29a13274076051351e701ed7f9392752adafe4f3499109170e976182f2d74

Observation 80bad068-16b6-43ad-80b8-bf7d4bd7d0d3 · inbound

Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving cites this paper.

Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:24:23.914142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T15:24:23.756052Z digest=sha256:241ef8990f10a841e337b2d5aca1447edf90751c3831a917a268d46e6dc00e82

Observation f69f2ee9-8c18-4e30-9e69-12438284e762 · inbound

CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving cites this paper.

CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:47:13.263203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:47:13.263203Z digest=sha256:137386abea52e48482e2009231003276a6f7b0e4ce483f6190a318cf4048b692

Observation c4908c9d-5c01-49b5-9620-2d62f93274a1 · inbound

SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels cites this paper.

SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:11:45.426681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:11:45.426681Z digest=sha256:a5f33b57700564081557f3d25b9f72fcfedb4acc7e4a64a79ccbf38f17bc80dc

Observation 096a0178-c3d0-405d-aea7-f3aec25945d9 · inbound

BEVCALIB: LiDAR-Camera Calibration via Geometry-Guided Bird's-Eye View Representations cites this paper.

BEVCALIB: LiDAR-Camera Calibration via Geometry-Guided Bird's-Eye View Representations BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:12:15.488257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:10:31.571968Z digest=sha256:51ac6a54925c6cac2640c27a9d66e7389547b6f8bc1543feae8e82b58a211a0d

Observation a6edca6f-c2c7-4d14-8e66-efec5e422ac1 · inbound

S2GO: Streaming Sparse Gaussian Occupancy Prediction cites this paper.

S2GO: Streaming Sparse Gaussian Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:09.658408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:09.658408Z digest=sha256:11635f6dad5f3a612dce547dcf6b9ea97be25cba6cba397d37a2422adcb9c194

Observation 717ab366-15ab-45aa-8d82-ac83504eb1f4 · inbound

FocalAD: Local Motion Planning for End-to-End Autonomous Driving cites this paper.

FocalAD: Local Motion Planning for End-to-End Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:17:15.800775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T10:15:53.988156Z digest=sha256:fcf21f8e8713cc779fefd27700542d454e74d44a6b6bbcf1fb61931bdba5b13c

Observation bd4ee98a-b979-4876-a424-03869bb94fc0 · inbound

Learning to Generate Vectorized Maps at Intersections with Multiple Roadside Cameras cites this paper.

Learning to Generate Vectorized Maps at Intersections with Multiple Roadside Cameras BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:22.884613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:22.884613Z digest=sha256:135853a211e09a773721175bd396da6c9594ad36b05b9613db00c8d72f3f30df

Observation aa3b2e12-822b-4e27-a50c-77f6bfcf1f42 · inbound

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection cites this paper.

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:42:00.957398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:42:00.957398Z digest=sha256:f630cf248dc101ba74b532264b7c15185cb81e5ec174915f2accc7e92e9aebfd

Observation d418e1cc-9c3e-4074-99fe-07da8eb6ef48 · inbound

Progressive Bird's Eye View Perception for Safety-Critical Autonomous Driving: A Comprehensive Survey cites this paper.

Progressive Bird's Eye View Perception for Safety-Critical Autonomous Driving: A Comprehensive Survey BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-05T22:06:55.337304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:06:55.337304Z digest=sha256:c34283fdf68a9bdc88a1e8c45dce7b2316e8688f6fd3b518be074cd5b5f4a2e1

Observation 45ada445-1397-4301-9130-0e54f13ea8f8 · inbound

SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer cites this paper.

SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T14:55:48.356558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:55:48.356558Z digest=sha256:1ea0e47558b4ef58e89afc2c37a1ff3797dc41ef105e1a72ba0ffb5006aaef81

Observation 0b72d4ed-8679-4e5f-a2b2-d9c4406ad130 · inbound

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models cites this paper.

Vehicle-to-Infrastructure Collaborative Spatial Perception via Multimodal Large Language Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:42:47.959381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:42:47.959381Z digest=sha256:ab874639a18192b55c6c925f60bebb00c325f79bd0721699cda1a2127c916f1b

Observation 6ded0fc4-1f91-46a2-8582-bb5106ccb31b · inbound

Semantic Causality-Aware Vision-Based 3D Occupancy Prediction cites this paper.

Semantic Causality-Aware Vision-Based 3D Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:48:57.622231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:48:57.622231Z digest=sha256:8b0089b97ced6910e0a4b82ba5d05bc38aa77db7df29688861a052cfa3f64990

Observation 73dae037-6dbb-4b3e-83b8-84eb82da462a · inbound

DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion cites this paper.

DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T09:29:20.708420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:29:20.708420Z digest=sha256:4622180cf791a6af4ed78536cde9318fe21d293bb64dada8b7ddd51d762a46d7

Observation f2ee8ce3-f025-44bc-a2cf-8780d356ba4e · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.632850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:8e378ffa33e40e63dafd90b3a1cd8d69f8c9f1230476c5c1c3f7296e7171f982

Observation 4637504e-1921-45fc-ad75-5ca91801d746 · inbound

Fast-BEV++: Fast by Algorithm, Deployable by Design cites this paper.

Fast-BEV++: Fast by Algorithm, Deployable by Design BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T18:44:18.886433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T18:42:37.182403Z digest=sha256:c121e048fe066adc651a0f4daeec76bd57d3f8645628be5679de8db2175f70c8

Observation 7fddc862-931f-4209-9130-5905bd0d3039 · inbound

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving cites this paper.

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T12:06:27.585778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:06:27.585778Z digest=sha256:cb02297e06b97518a5f9084091d0b1b238e660778df0f734593bd838ca54fcaa

Observation f73c6746-4829-4ace-8259-f46256c95782 · inbound

More than the Sum: Panorama-Language Models for Adverse Omni-Scenes cites this paper.

More than the Sum: Panorama-Language Models for Adverse Omni-Scenes BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:00:02.865442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T13:58:18.635091Z digest=sha256:9d211a8134c27bb110c4f7cce5856fb1495b43062c36e6a036cae1bfbd1c3294

Observation b4e21461-f0d2-488e-9dbd-eb891e87f78c · inbound

BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving cites this paper.

BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:03:12.414025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:00:21.130075Z digest=sha256:c6f8353611ae79aa1da77d6c0b2636de96b3d4bc3199bbe537ee598dae6ce0bc

Observation 442fffd0-908d-4b5e-b539-94bf4942ff64 · inbound

Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving cites this paper.

Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:00:56.801261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:43:52.542973Z digest=sha256:5d5c5f31b3ea212c103f16c78b60f0b8cf5bdf0fecc0e1207f641c1ffef3aa94

Observation 5366bd63-cc7f-4830-be9a-7fa8f7bb297f · inbound

Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion cites this paper.

Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:47.954582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T20:22:04.663359Z digest=sha256:19c38c25cbe501c49295c9ad0fce052c73b4dd5fa128c9464a3ad3ce43d5efe5

Observation 5fa18e0b-93cc-49c5-ac98-343690cc4eb2 · inbound

ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution cites this paper.

ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:56:14.121985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T15:58:59.744568Z digest=sha256:2d49c5fbeb370cde9f40d2cb5ed430a8076ad8d0e7964454c45958b0e37bce19

Observation c8d8e482-4ed2-4c66-a9c4-9054d4a82e92 · inbound

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making cites this paper.

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.053667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T17:01:31.011188Z digest=sha256:edae6053f5863a2150e9420dfcfa86cca65af431118fde6d2fd661b7d2a2fc24

Observation 2d7ede07-3691-4f8c-8848-21f70c4366bd · inbound

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding cites this paper.

SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:26.590225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T00:53:35.188721Z digest=sha256:71b4c7fca45f131856166bacc6e1b51cb79cfe0cf8b5c3e28fb60f5d46f1298a

Observation ef25aee8-786a-4d1c-80fe-cf3bee87bdf5 · inbound

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers cites this paper.

CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:49:44.370206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T04:47:32.476614Z digest=sha256:5a9a0d1a0f4a9f4d5f08929225bb28bd9ba3280264382935445c3cf574bd077a

Observation 21f924e8-64f1-450b-98a1-7a4e0013d1ea · inbound

Deformba: Vision State Space Model with Adaptive State Fusion cites this paper.

Deformba: Vision State Space Model with Adaptive State Fusion BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:03:57.771091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T05:03:12.736238Z digest=sha256:a2655ef8ba8c500b9b529fb94a7bca6d4226f0e53dbd68838a962fa2b1356fc4

Observation 025a8355-0606-4108-856f-3dbb2e19596a · inbound

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation cites this paper.

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.452299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T11:11:46.428727Z digest=sha256:4e337c5fc57b0e0c6d8b5dee1afc049b22c3f184e9b31894f4d97b3b3174a18b

Observation ac39401a-4851-47df-b270-f419b7319e87 · inbound

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation cites this paper.

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T12:34:16.033188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:34:16.033188Z digest=sha256:26db54e6041c80f594ac46e94b5339a339c0d5e37399e990eee2e99110b26dcb

Observation 11892c46-f107-46b5-9eda-0a28e3b81e72 · inbound

Isolation-aware Scheduling Framework for DNN-based End-to-End Autonomous Driving System on Tile-based Accelerators cites this paper.

Isolation-aware Scheduling Framework for DNN-based End-to-End Autonomous Driving System on Tile-based Accelerators BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:47:44.612320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T11:49:46.697087Z digest=sha256:ff74b5edb8eb493765a635e8b20e50a2af8e5ef5daf48f15ecef4c94ccad8c81

Observation 6f9966dc-0deb-4187-9a50-bc27b622f84c · inbound

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models cites this paper.

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:22.585056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T07:08:41.071757Z digest=sha256:ae2fc8972d7f77b6caf5db76c757caec292044d47ceb39538bb9a8c8bc83c9cb

Observation 140ef37a-73d5-423c-b6f5-1f951345f183 · inbound

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models cites this paper.

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-15T10:49:28.959330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:49:28.959330Z digest=sha256:ba9cabcdabee939b40350c42dff826c7e89b7b0297e54f518cdf32a8e1df8060

Observation b4c6a4db-587d-4420-9252-6586c38d9d2d · inbound

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling cites this paper.

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.659347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T04:35:58.372801Z digest=sha256:fe2bd26822bd6d14cb8d5896f7106f9df9341cdbaacd33e15813b8730dbc0681

Observation 01b31408-dafb-43ed-838e-739b059f9d83 · inbound

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling cites this paper.

Panoramic Scene Understanding: A Survey from Distortion-Aware Engineering to Sphere-Native Modeling BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T09:58:51.929181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:58:51.929181Z digest=sha256:d9aef32612eb7948eb53b46a0473f32e2f13df03b8787333edc510fbaec431f0

Observation 2657712e-93b6-4f30-94bc-52cbdd4fbd41 · inbound

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction cites this paper.

FDR-Occ: Factorized Dense Routing for Full-Spectrum 3D Occupancy Prediction BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T23:42:57.473601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:42:57.473601Z digest=sha256:c5b08f43cdf077216314dbacfa839191e7302e3d67ee3041d72bd48b78d633ce

Observation b06769f7-8c62-460c-94c6-aad4a59852fa · inbound

RayOcc: Occlusion-Aware Ray Occupancy Estimation via Gaussian Mixture Intensity cites this paper.

RayOcc: Occlusion-Aware Ray Occupancy Estimation via Gaussian Mixture Intensity BEVFormer: Learning Bird's-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T17:26:12.201838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:26:12.201838Z digest=sha256:f27101576eef0571addf89d9819a465f7af739f58dd447b4812d25f78cf2fef2