Pith. sign in

Paper Citation Record · LEDGER

Physical Informed Driving World Model

As of 23 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 6 inbound Pith citation observations for arXiv:2412.08410.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08410 v2

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:58:58.492878Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:31:52.984933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T17:51:54.796433Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b85c0c5-43be-41c7-b032-495d5ca1f6dd · outbound

This paper cites Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation.

Physical Informed Driving World Model Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.286202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.286202Z digest=sha256:853e29a5a17fbbbd79df138e21c95d5c4b0a87d2fdc02a2355f4f35c74829e6b

Observation 75343fff-0023-4f32-94c7-0de5bcc45895 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Physical Informed Driving World Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.291656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.291656Z digest=sha256:dbfb707884fe7fc8bb2afa9a5d659cfafb1153c57fcef0f75a1d80e4246d0309

Observation 5ea56432-d879-4eff-9285-8c966f4918e2 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

Physical Informed Driving World Model Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.124634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.296448Z digest=sha256:b5671d9226c498f49446336ce817518784b979001bf2c0dd8857b35691bb2dac

Observation 987239e9-29b5-4c09-826a-43275ec99950 · outbound

This paper cites Virtual KITTI 2.

Physical Informed Driving World Model Virtual KITTI 2

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.301381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.301381Z digest=sha256:2adf47edee9b71469315020cb2a59c260e1c7076595d4c6ad090dddbdd3b5338

Observation a0521e8f-14e0-49c9-8f01-6c981d7c9f8d · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

Physical Informed Driving World Model nuscenes: A multi- modal dataset for autonomous driving

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.111242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.306570Z digest=sha256:876e4aba4827f0a715f6820a84817b28eb399b842d5f79d1e8bb5bfa59e9c37a

Observation 1d4b8f78-ef90-42aa-bc7f-c1477707f121 · outbound

This paper cites CARLA: An open urban driving simulator.

Physical Informed Driving World Model CARLA: An open urban driving simulator

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.097512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.311038Z digest=sha256:1343e01992fd04c97a307a422eece191486ff948753d771840618c2b5f1217d4

Observation b98e65e1-a64d-4f98-b560-91243918a682 · outbound

This paper cites Magicdrive: Street view generation with diverse 3d geometry control.

Physical Informed Driving World Model Magicdrive: Street view generation with diverse 3d geometry control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.082815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.315655Z digest=sha256:fa30403d4e13d9ed0e5dc0d80e67b79631d6f2be9c284b3f73bd628ce63edcb4

Observation 03fda73b-3a78-451d-aa1d-89884b150742 · outbound

This paper cites Gan-based virtual- to-real image translation for urban scene semantic segmen- tation.

Physical Informed Driving World Model Gan-based virtual- to-real image translation for urban scene semantic segmen- tation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.069599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.319643Z digest=sha256:fa8fd8a247c14b590795757669010e6d7c271af2257a43f2007036d357dcceb5

Observation f02bec2d-0df7-47cb-93b1-e0cb677f3d8b · outbound

This paper cites Learning video rep- resentations of human motion from synthetic data.

Physical Informed Driving World Model Learning video rep- resentations of human motion from synthetic data

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.055054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.324204Z digest=sha256:709f28d4377ca61fcc52e5c387692af1b2f2aee1cddc98392dec63962b0d3f95

Observation 18ab6450-a9f5-4acb-9073-c03814939a8e · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Physical Informed Driving World Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.328661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.328661Z digest=sha256:0a61351b5cd3aa0f6dd51f2cacbc71c0ff0a1dbb3872021f3e8fa3e325dd97f8

Observation ab10fe61-4f49-4b5e-bdc9-be30527bf3b3 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Physical Informed Driving World Model Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.332946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.332946Z digest=sha256:b10e614851ede4d200f01463c78bc9f20e12ed9bbd8f48f08e867690c8382bce

Observation 570188ec-28e4-48cd-a558-2dc92374c107 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

Physical Informed Driving World Model Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.042253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.337261Z digest=sha256:ad16ba37f6585c26255a0f73a25954eabe34f9866f61954f9863b7eb446c4b67

Observation f08eac2d-9bd4-47e3-9412-e15a8629f442 · outbound

This paper cites Classifier-free diffusion guidance.

Physical Informed Driving World Model Classifier-free diffusion guidance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.028635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.341410Z digest=sha256:d8b20a0febdcfc2290325a792d137659275d42fe9056c9652b2b8ef709f213bb

Observation 1fd1a30d-f4a4-4914-a1ab-7d9d60b7dff3 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Physical Informed Driving World Model Denoising diffu- sion probabilistic models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.345291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.345291Z digest=sha256:443a153287dd61063048b859bae3aee50a1108a91aa56e79255316d0aac33704

Observation cf78aa91-1a4e-4d1e-9d16-020a886fe525 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Physical Informed Driving World Model Imagen Video: High Definition Video Generation with Diffusion Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.349163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.349163Z digest=sha256:d98c24c0ef35859f57a84b609044755fc040b37decb591a63db076f66a6dd2e6

Observation e682f2e9-29f4-4a09-b41d-33420f9fc021 · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

Physical Informed Driving World Model ADriver-I: A General World Model for Autonomous Driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.353543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.353543Z digest=sha256:17d5a372713c6811f32273130e01c998c60b5069617c5bbf026b1f39f7a5332c

Observation 3732d6b9-4e17-47fa-b897-4dabdc7a8d15 · outbound

This paper cites Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023.

Physical Informed Driving World Model Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.007453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.357846Z digest=sha256:b715df2bfad807130aba7d2971d9dc268c0cbc37a1a7c54b6c5665b68c6e6f4f

Observation 756cf46a-6078-4185-9a99-6265064ede36 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Physical Informed Driving World Model Gligen: Open-set grounded text-to-image generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.994439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.361557Z digest=sha256:8b2d8d42930338d94c1ce40b1b683afaf417426da208bc7a22d62dee068709e5

Observation 935c05c4-bc57-4ca6-ae59-9fc8f3096eb2 · outbound

This paper cites Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024.

Physical Informed Driving World Model Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.980598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.365871Z digest=sha256:f3593d4e20537b510c9ea91ba76467f1a21f920af2d3bba9b9fdefc4f6db9954

Observation fc67a0b2-7b4b-4fc6-8331-72b88f1c2bf3 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

Physical Informed Driving World Model Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.967463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.369874Z digest=sha256:92511205e5bfd360b130d4877635484aa44e5ef647c5ccb0ddb654a9b4fdb17b

Observation fe44afa6-d804-46f2-adda-0e2d06dfbf58 · outbound

This paper cites Scalable diffusion models with transformers.

Physical Informed Driving World Model Scalable diffusion models with transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.373643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.373643Z digest=sha256:19a15eb56766aaf168c75d467889aaddeeff4d740e060cd6511859871b57e38d

Observation 7c6fbc32-7cfd-4b7d-9b09-51467ffa4ebb · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Physical Informed Driving World Model SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.377429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.377429Z digest=sha256:6e06d86ef8848c41ec2e9850962f584bf287ddeee8b4df593be178a36da49297

Observation 000271a3-7cdf-48c8-8fd4-d4982ba29049 · outbound

This paper cites Exploring the limits of transfer learning with a uni- fied text-to-text transformer.

Physical Informed Driving World Model Exploring the limits of transfer learning with a uni- fied text-to-text transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.945335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.381250Z digest=sha256:8e17ce51a233cf83deb98c8f0a56aacba6a82af065ef8b58d754f84590147ac4

Observation 91ff9ac6-cfd6-4cd7-b9cd-09546e831195 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Physical Informed Driving World Model Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.384644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.384644Z digest=sha256:1f0f6f8100bf8f7cd9678af8e72f0f078c2aadb0eba15e72428c4e65e6d015f3

Observation a36f5579-6f58-457f-a3e9-6b94fb7e4cb1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Physical Informed Driving World Model High-resolution image synthesis with latent diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.388465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.388465Z digest=sha256:38f625917148fb2bf9be0601363ccc981411a5bea7de18a4927ab08a933b6fde

Observation 148fed5e-2a00-4fef-994d-477f64264158 · outbound

This paper cites The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes.

Physical Informed Driving World Model The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.924702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.391900Z digest=sha256:1cdf3fed3bfeb5bb6fdcd13c760c1776d33c4d5d0df7516710bb2aae1daa6214

Observation 4482455a-6548-4333-9044-ee7eb77ee734 · outbound

This paper cites Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016.

Physical Informed Driving World Model Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.911860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.396122Z digest=sha256:1b5bc78eb7d6185fe399ce15ab9ff287208a85cd74314cefd261f7681008a959

Observation ba0116da-ad7c-4838-918a-474525c11b6a · outbound

This paper cites Learning 12 from simulated and unsupervised images through adversarial training.

Physical Informed Driving World Model Learning 12 from simulated and unsupervised images through adversarial training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.899709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.399608Z digest=sha256:2675b051fc34d4826c5fc54b55c2802186543e4c2d4044c2ee7d294687604b8d

Observation 29c9a3a0-48ed-415d-b8bb-71af4b71e4b5 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Physical Informed Driving World Model Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.403342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.403342Z digest=sha256:93d3b40e34e0167bdc17ee3262584cb7974f7f4a9717856e02a25c6b3d28ab86

Observation f4eafede-498a-4c98-ab1a-c614d33a7628 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Physical Informed Driving World Model Score-based generative modeling through stochastic differential equa- tions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.407998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.407998Z digest=sha256:06a82e7f36b4f5624e65319d74afa637c5defbcb8944363b3a512dbea577b63e

Observation 76aa1697-7fa0-4fbc-a1c7-4b5ac5bd28c3 · outbound

This paper cites Street-View Image Generation from a Bird's-Eye View Layout.

Physical Informed Driving World Model Street-View Image Generation from a Bird's-Eye View Layout

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.411831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.411831Z digest=sha256:3dffbd9f002fb8e183efec2e44a8fab9951800a28144745437df31a037c326d6

Observation 043a437b-8ece-492d-9f88-e875ab79b733 · outbound

This paper cites MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion.

Physical Informed Driving World Model MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.416036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.416036Z digest=sha256:45bf85a088f6f835ddfe56350ef3fe78274bae1d7f631280deda8d77db56e72a

Observation 90dfeb38-14e8-42b5-8ec6-0ab6ae84e8cf · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Physical Informed Driving World Model Domain randomization for transferring deep neural networks from simulation to the real world

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.419800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.419800Z digest=sha256:d56aa56e5b5ac48b93088b98b89545fabe45d023c996195b9816bd7874f4a88d

Observation 00f8a4eb-a254-474f-847c-7390af8597c4 · outbound

This paper cites Consistent view synthesis with pose-guided diffusion models.

Physical Informed Driving World Model Consistent view synthesis with pose-guided diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.871909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.423768Z digest=sha256:3bd2fcdefb8a99010e1a961ba353fb10f16c74227ecf2027ffb45435e7a6d629

Observation 9fd47125-8eb4-4303-9e66-2585f4993afe · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Physical Informed Driving World Model Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.427881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.427881Z digest=sha256:dbc9c1f95cc135953fa81c27ef3ee5aac14a02d6ebc41f44897876d996cc7573

Observation 49bfcf36-bfed-4090-9e16-ab534fe517e0 · outbound

This paper cites Attention is all you need.

Physical Informed Driving World Model Attention is all you need

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.859969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.432062Z digest=sha256:d15fa1c2f3ab11c6352d23a5578f6512006adc75ae9e5007ed1037a0a66bda8b

Observation 4e491b20-097c-4473-b511-1ea4fa851230 · outbound

This paper cites Exploring object-centric temporal modeling for efficient multi-view 3d object detection.

Physical Informed Driving World Model Exploring object-centric temporal modeling for efficient multi-view 3d object detection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.847896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.436030Z digest=sha256:5ce74ca0eae18b0cb5b0e47119470cd9d7c4bc54d3bc77b6d0eddf12e5b57ceb

Observation 430b276b-b838-4d2e-85ad-6320e44c9131 · outbound

This paper cites Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023.

Physical Informed Driving World Model Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.834623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.439999Z digest=sha256:2d5e723dd455e09bd57c5691d7c8f7ddf159d3d342977ba69683b0ee45c35fdc

Observation 892263d8-6b64-47c9-a70a-6e8de63377f7 · outbound

This paper cites DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving.

Physical Informed Driving World Model DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.444090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.444090Z digest=sha256:2fdf8cf2ad948385c1657ff130eeaa5c390c4d24d6fdd5f9faeb0ec08d0dea6d

Observation d2a94574-c5cc-4af1-8eea-8bde469010f2 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023.

Physical Informed Driving World Model Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.821545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.447595Z digest=sha256:edb384661b02d05e1ca7c5eebbf2eda04b191d43379a4c2b9c6433ab703d344b

Observation b72f2ca1-c71f-4fd6-923e-f1fbfb03b7b2 · outbound

This paper cites Panacea: Panoramic and controllable video generation for autonomous driving.

Physical Informed Driving World Model Panacea: Panoramic and controllable video generation for autonomous driving

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.808701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.451136Z digest=sha256:bf7c541adb1985d22f143f376b152bcb2d512401004ca639d2022cbdeea19115

Observation 9d3a015a-af50-402d-b044-4a7af83a78b3 · outbound

This paper cites BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout.

Physical Informed Driving World Model BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.455339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.455339Z digest=sha256:b67991b11acf70b927a12a4c6d1a17dfce56e524f02840f7cc46912ecfa90ef5

Observation eb4378b4-cb85-42ff-b4b1-7f2dbfaecf36 · outbound

This paper cites Magvit: Masked generative video transformer.

Physical Informed Driving World Model Magvit: Masked generative video transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.796348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.459666Z digest=sha256:c6cacc5a91bc65b977b68f72316f05ad5df8970366ce5bf2e43cb35c454c69d2

Observation 88a19d93-3bd4-4807-9d47-24f40d1aadb2 · outbound

This paper cites Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024.

Physical Informed Driving World Model Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.783763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.463513Z digest=sha256:7862c5f6d714c5d5f254e4e4f17fd5dd92fb9c8db8a40aa59e6d0ee13fbbe75c

Observation 85472ba3-cd49-4454-86d6-6f450f65df25 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Physical Informed Driving World Model Adding conditional control to text-to-image diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.467551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.467551Z digest=sha256:84a606d655d96b2c6b5e705da3f3d43726da62472e821c2ed4909dd573f6b9b0

Observation d760a2ca-8a8a-463e-999a-4977e19d6fc3 · outbound

This paper cites Collaborative and adversarial network for unsupervised do- main adaptation.

Physical Informed Driving World Model Collaborative and adversarial network for unsupervised do- main adaptation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.764198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.471434Z digest=sha256:b346a32e43569ebf7584067523dcf2d0525163f0aa2c0260a4469b2224181dbb

Observation 4b2863e2-328f-4295-9fad-e79343f57c3b · outbound

This paper cites DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation.

Physical Informed Driving World Model DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.475852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.475852Z digest=sha256:38b04b67dc7efc6116fb0770d2b08ed58f448b0f58f737facf2d5c449220c2b1

Observation 3768ceed-6242-4fd5-851f-519b592fb326 · outbound

This paper cites GenAD: Generative End-to-End Autonomous Driving.

Physical Informed Driving World Model GenAD: Generative End-to-End Autonomous Driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.480102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.480102Z digest=sha256:aec938544b3678242a2460b312f2204ca9601ebeb131b6f64e045ebaa2774243

Observation bd638ff6-3a72-44aa-a7ac-bd090241bf8c · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

Physical Informed Driving World Model Open-sora: Democratizing efficient video production for all, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.752042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.484128Z digest=sha256:7053accae1573e40761d18a27bb7baf9842209bd0c4432cfab103d8167415e16

Observation 16d0aac3-1ead-4981-b815-6823e7a80c92 · outbound

This paper cites patchified.

Physical Informed Driving World Model patchified

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.738519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.488155Z digest=sha256:3e61115e24783fb07f330de1b642a8efca191db79166be387b3e9f1c035f2200

Observation 35f79062-b189-47a1-960b-f48e0bd19b09 · outbound

This paper cites Simulation-to-Real Visual Translation.

Physical Informed Driving World Model Simulation-to-Real Visual Translation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.724458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T17:58:58.492878Z digest=sha256:da82d3861daa11a7afa1a2588e73c0ef0762010a2a484e859c19d9554f8721cd

Pith citing papers

Observation 41ebb73c-9c33-49a7-98e5-dd8f28167298 · inbound

A Survey of World Models for Autonomous Driving cites this paper.

A Survey of World Models for Autonomous Driving Physical Informed Driving World Model

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-10T18:31:52.984933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:31:52.984933Z digest=sha256:1b0321355d309f0cb9d1f33aca849afcad52e7a1c0f671d378189db0f4c29a15

Observation 3afa36f0-f4de-441e-aba1-ea4e330d1edf · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment Physical Informed Driving World Model

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.799714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:d13a3c3d5477601a204b9f51fcd427d1da7cf83f74f7411e9dfa5c10bd8beec5

Observation 24f0798c-92af-4e58-bc33-13b694187851 · inbound

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities cites this paper.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Physical Informed Driving World Model

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.587203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.587203Z digest=sha256:4392433bcea2e2766a28b8dccc94a430b6068533ff7b880a95deff0e3469bcd7

Observation 2c5cbbd6-3523-43cc-a13a-25dbb91d6b98 · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI Physical Informed Driving World Model

Reference 214

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:54.374022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:54.374022Z digest=sha256:2838fe8a8877f3dd2c98409241c98e287a123c6eb93c7afb3367070e51a92d5d

Observation a03e8934-99c5-43f2-bf1e-d4066f579245 · inbound

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World cites this paper.

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World Physical Informed Driving World Model

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-03T17:02:42.473873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:02:42.473873Z digest=sha256:212a896dbef7204d1e771dfdc78369c92aa38766e9e13564f77a7fc53dcc4c96

Observation dabfd1d9-faf5-4c98-aa6b-3507fe97cfca · inbound

MultiWorld: Scalable Multi-Agent Multi-View Video World Models cites this paper.

MultiWorld: Scalable Multi-Agent Multi-View Video World Models Physical Informed Driving World Model

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:09:08.093011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T05:06:11.514186Z digest=sha256:38f0d60b04f0cc18810aeecdaf4f3ace9a16861df2478cc4164ff0e05caade35