Pith. sign in

Paper Citation Record · LEDGER

Physical Informed Driving World Model

As of 16 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 6 inbound Pith citation observations for arXiv:2412.08410.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08410 v2

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:58:58.492878Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:31:52.984933Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T17:51:54.796433Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b85c0c5-43be-41c7-b032-495d5ca1f6dd · outbound

This paper cites Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation.

Physical Informed Driving World Model Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.286202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.286202Z digest=sha256:9ecbaf4dfdbf189783be0e0fcd024afcb533d7bde8ff810bcea62c86492f0c6c

Observation 75343fff-0023-4f32-94c7-0de5bcc45895 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Physical Informed Driving World Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.291656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.291656Z digest=sha256:baefaaf4fbe740a6b6318355c003320b322eccbcd35074b8063102caca118b8d

Observation 5ea56432-d879-4eff-9285-8c966f4918e2 · outbound

This paper cites Align your latents: High-resolution video synthesis with la- tent diffusion models.

Physical Informed Driving World Model Align your latents: High-resolution video synthesis with la- tent diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.124634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.296448Z digest=sha256:5d5f88950498dbefde4451c991d7df251905baf0071b055798f26f5701309863

Observation 987239e9-29b5-4c09-826a-43275ec99950 · outbound

This paper cites Virtual KITTI 2.

Physical Informed Driving World Model Virtual KITTI 2

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.301381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.301381Z digest=sha256:abdbb6d5e6d4fda48745a9c5f297104db69db5cb32f3baf5ad0603ffcca2190e

Observation a0521e8f-14e0-49c9-8f01-6c981d7c9f8d · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

Physical Informed Driving World Model nuscenes: A multi- modal dataset for autonomous driving

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.111242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.306570Z digest=sha256:9474951e3b7745b48fe97f9896b1effe117720b443ae0017a9844bf41c5edaae

Observation 1d4b8f78-ef90-42aa-bc7f-c1477707f121 · outbound

This paper cites CARLA: An open urban driving simulator.

Physical Informed Driving World Model CARLA: An open urban driving simulator

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.097512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.311038Z digest=sha256:064e6fdc227e7ae3d7fe4a0af8dca5a807fd006fa1870f9cdff707f0c58b26f2

Observation b98e65e1-a64d-4f98-b560-91243918a682 · outbound

This paper cites Magicdrive: Street view generation with diverse 3d geometry control.

Physical Informed Driving World Model Magicdrive: Street view generation with diverse 3d geometry control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.082815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.315655Z digest=sha256:dec37e733aebf1d01231ec65918b4f61c7207e6702cc406b264cfe917dfe2ca7

Observation 03fda73b-3a78-451d-aa1d-89884b150742 · outbound

This paper cites Gan-based virtual- to-real image translation for urban scene semantic segmen- tation.

Physical Informed Driving World Model Gan-based virtual- to-real image translation for urban scene semantic segmen- tation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.069599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.319643Z digest=sha256:cbec2e0880c8f1d83e2280d9b81d9470c1b02e5c4f8b378f2a366b8dc03f5d3d

Observation f02bec2d-0df7-47cb-93b1-e0cb677f3d8b · outbound

This paper cites Learning video rep- resentations of human motion from synthetic data.

Physical Informed Driving World Model Learning video rep- resentations of human motion from synthetic data

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.055054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.324204Z digest=sha256:45c961b1f7f82c49af941c81fd72d175f62e0e1b449d3fc2e8563be986e5dfa9

Observation 18ab6450-a9f5-4acb-9073-c03814939a8e · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Physical Informed Driving World Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.328661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.328661Z digest=sha256:18dffd06a30a3a90331845ebe77cf447dbda730c72d797856f5c45237d31cc66

Observation ab10fe61-4f49-4b5e-bdc9-be30527bf3b3 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Physical Informed Driving World Model Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.332946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.332946Z digest=sha256:c1f439c585c1bf189fe12bd5ebb43c687d5719632fddaeec2de43d365c5e6556

Observation 570188ec-28e4-48cd-a558-2dc92374c107 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

Physical Informed Driving World Model Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.042253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.337261Z digest=sha256:67e7c1ddaf105f501e4ba7571936e869c55fa2df968a314b069b73baf9b3080c

Observation f08eac2d-9bd4-47e3-9412-e15a8629f442 · outbound

This paper cites Classifier-free diffusion guidance.

Physical Informed Driving World Model Classifier-free diffusion guidance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.028635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.341410Z digest=sha256:cb446df389fe98dc9f0acaec00f9e0f2e510a88aa80162c175dc1070f89207f0

Observation 1fd1a30d-f4a4-4914-a1ab-7d9d60b7dff3 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Physical Informed Driving World Model Denoising diffu- sion probabilistic models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.345291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.345291Z digest=sha256:84ef4cbfea5cf1684c2bc664fafabe0e2fd30314130571234c930f0cdb215c3c

Observation cf78aa91-1a4e-4d1e-9d16-020a886fe525 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Physical Informed Driving World Model Imagen Video: High Definition Video Generation with Diffusion Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.349163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.349163Z digest=sha256:af6b6aa6924e4b181607f192397c123372529d53f466977f1e569a4836424435

Observation e682f2e9-29f4-4a09-b41d-33420f9fc021 · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

Physical Informed Driving World Model ADriver-I: A General World Model for Autonomous Driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.353543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.353543Z digest=sha256:8d2ed162209a08c272ab03028a0bdc33b5bdf9a4c81a53a63e26dfc6d47a7ea9

Observation 3732d6b9-4e17-47fa-b897-4dabdc7a8d15 · outbound

This paper cites Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023.

Physical Informed Driving World Model Drivingdiffu- sion: Layout-guided multi-view driving scene video gener- ation with latent diffusion model, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:59.007453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.357846Z digest=sha256:5114421d73d9db2c4af8effb8408279f82b4d56ef58f4b12affb1eb6846977b5

Observation 756cf46a-6078-4185-9a99-6265064ede36 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Physical Informed Driving World Model Gligen: Open-set grounded text-to-image generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.994439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.361557Z digest=sha256:69765916ecbb8725e6bb71aed986e03a73f32fbc71adcc1bb19874a4ebfdfc7a

Observation 935c05c4-bc57-4ca6-ae59-9fc8f3096eb2 · outbound

This paper cites Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024.

Physical Informed Driving World Model Wovogen: World volume-aware diffusion for con- trollable multi-camera driving scene generation, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.980598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.365871Z digest=sha256:41d173de987038673a81add958c6282ab395d15ff17369b6af12292a8f7ec572

Observation fc67a0b2-7b4b-4fc6-8331-72b88f1c2bf3 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

Physical Informed Driving World Model Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.967463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.369874Z digest=sha256:f6cb29a75d2bc1e72a2bd4c672c034e534e8228b25beaae2c1698826a76344b0

Observation fe44afa6-d804-46f2-adda-0e2d06dfbf58 · outbound

This paper cites Scalable diffusion models with transformers.

Physical Informed Driving World Model Scalable diffusion models with transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.373643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.373643Z digest=sha256:c9d8fe3ae9507b824d3652ae4dcdb923515b34892fae183649c37a3aa8bfddd8

Observation 7c6fbc32-7cfd-4b7d-9b09-51467ffa4ebb · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Physical Informed Driving World Model SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.377429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.377429Z digest=sha256:b360d7b810b066ee463aa1c992dd16a8ccfd4400dcf983ccbde54b301a672424

Observation 000271a3-7cdf-48c8-8fd4-d4982ba29049 · outbound

This paper cites Exploring the limits of transfer learning with a uni- fied text-to-text transformer.

Physical Informed Driving World Model Exploring the limits of transfer learning with a uni- fied text-to-text transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.945335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.381250Z digest=sha256:6e8f3d3ed78c1fbda0f2eb9ae73895dd1285148bec57bbb533b9cd04423ffaf3

Observation 91ff9ac6-cfd6-4cd7-b9cd-09546e831195 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Physical Informed Driving World Model Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.384644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.384644Z digest=sha256:2aebf2dc9bb82913860b7d4c52a9e2917d83bc776cd0dcce518343598a17426b

Observation a36f5579-6f58-457f-a3e9-6b94fb7e4cb1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Physical Informed Driving World Model High-resolution image synthesis with latent diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.388465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.388465Z digest=sha256:a17d9291609f45c3c82a7cafdb68ee27d154c92fec0f600ceef78c869f87c68f

Observation 148fed5e-2a00-4fef-994d-477f64264158 · outbound

This paper cites The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes.

Physical Informed Driving World Model The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.924702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.391900Z digest=sha256:fb68ebc52eabce6303dbf52a2fc46d604a17e770ef876724cca87129ca14bf88

Observation 4482455a-6548-4333-9044-ee7eb77ee734 · outbound

This paper cites Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016.

Physical Informed Driving World Model Uniad: A unified ad hoc data processing system.ACM Trans- actions on Database Systems (TODS), 42(1):1–42, 2016

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.911860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.396122Z digest=sha256:d43e691316fc43b0b4e7888d2a13e1ed141b40472f470246c40d612ec4bdc67f

Observation ba0116da-ad7c-4838-918a-474525c11b6a · outbound

This paper cites Learning 12 from simulated and unsupervised images through adversarial training.

Physical Informed Driving World Model Learning 12 from simulated and unsupervised images through adversarial training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.899709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.399608Z digest=sha256:d9ba8a4102173768493415d8d1d1fa61936c2a506c5ea0431a02f15c69413ead

Observation 29c9a3a0-48ed-415d-b8bb-71af4b71e4b5 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Physical Informed Driving World Model Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.403342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.403342Z digest=sha256:17712a64658fd33447b428371f827e725f4261842473ca4a961e1463948b6f0e

Observation f4eafede-498a-4c98-ab1a-c614d33a7628 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Physical Informed Driving World Model Score-based generative modeling through stochastic differential equa- tions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.407998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.407998Z digest=sha256:ac8a10f7fc847e0dda289369a8fd3625edd1696e9dc71f76b562b49ed7400c25

Observation 76aa1697-7fa0-4fbc-a1c7-4b5ac5bd28c3 · outbound

This paper cites Street-View Image Generation from a Bird's-Eye View Layout.

Physical Informed Driving World Model Street-View Image Generation from a Bird's-Eye View Layout

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.411831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.411831Z digest=sha256:5247f1919ef7e081e4b6fc9194cc9bca0e7107eab018f398c3e3b1876a762254

Observation 043a437b-8ece-492d-9f88-e875ab79b733 · outbound

This paper cites MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion.

Physical Informed Driving World Model MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.416036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.416036Z digest=sha256:dee01119f3d7e5cfbdb29289d26b859117c9aba55014bdf35f2aaa6d3b76fb94

Observation 90dfeb38-14e8-42b5-8ec6-0ab6ae84e8cf · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Physical Informed Driving World Model Domain randomization for transferring deep neural networks from simulation to the real world

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.419800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.419800Z digest=sha256:18588b5018a9c51077442170968d0f4540193183ca5af77f5c9764e249991851

Observation 00f8a4eb-a254-474f-847c-7390af8597c4 · outbound

This paper cites Consistent view synthesis with pose-guided diffusion models.

Physical Informed Driving World Model Consistent view synthesis with pose-guided diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.871909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.423768Z digest=sha256:e5e4217b6bdc0192a0b12d6470ac3df636d2e72818734b833f16603473f8ffcc

Observation 9fd47125-8eb4-4303-9e66-2585f4993afe · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Physical Informed Driving World Model Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.427881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.427881Z digest=sha256:8c8e0a74cee0b3a79d72471bba8717c4f59f511119520fa8362618cda6123a73

Observation 49bfcf36-bfed-4090-9e16-ab534fe517e0 · outbound

This paper cites Attention is all you need.

Physical Informed Driving World Model Attention is all you need

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.859969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.432062Z digest=sha256:a951d9c725d6e4b386efe8bce671c6b58c0d03b57a490e62ad66d5852fd2fbe2

Observation 4e491b20-097c-4473-b511-1ea4fa851230 · outbound

This paper cites Exploring object-centric temporal modeling for efficient multi-view 3d object detection.

Physical Informed Driving World Model Exploring object-centric temporal modeling for efficient multi-view 3d object detection

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.847896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.436030Z digest=sha256:12ad07291059c7266fd993bea6811869667ff8eaec4a0ca2dfd19a04755f4138

Observation 430b276b-b838-4d2e-85ad-6320e44c9131 · outbound

This paper cites Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023.

Physical Informed Driving World Model Videofactory: Swap at- tention in spatiotemporal diffusions for text-to-video gener- ation, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.834623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.439999Z digest=sha256:36d483ddbe0f2739065d1cab4e4175db678ad621b92f40fcf0fb0f16d4371077

Observation 892263d8-6b64-47c9-a70a-6e8de63377f7 · outbound

This paper cites DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving.

Physical Informed Driving World Model DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.444090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.444090Z digest=sha256:7701adb3ca639b456f607537932ddf67d9cf5967a746be833022c10ec43049f7

Observation d2a94574-c5cc-4af1-8eea-8bde469010f2 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023.

Physical Informed Driving World Model Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.821545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.447595Z digest=sha256:1adb27d648e33e2d724839c2e5288a0a6e30afcafcf4bc5fe9ed5d57da81ad55

Observation b72f2ca1-c71f-4fd6-923e-f1fbfb03b7b2 · outbound

This paper cites Panacea: Panoramic and controllable video generation for autonomous driving.

Physical Informed Driving World Model Panacea: Panoramic and controllable video generation for autonomous driving

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.808701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.451136Z digest=sha256:c4b6ad43285ff8ab4f14ac9fde64e22a06e7404273e8871418c3d51dd09f661c

Observation 9d3a015a-af50-402d-b044-4a7af83a78b3 · outbound

This paper cites BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout.

Physical Informed Driving World Model BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.455339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.455339Z digest=sha256:0a90221f8f22ce09377213c128d7479680f004bea74de4d48b00bc69446063ce

Observation eb4378b4-cb85-42ff-b4b1-7f2dbfaecf36 · outbound

This paper cites Magvit: Masked generative video transformer.

Physical Informed Driving World Model Magvit: Masked generative video transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.796348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.459666Z digest=sha256:908b78a47d64bc83e50f0bd7a3995b9de034bdee14af15f7c5b95fe5576a801c

Observation 88a19d93-3bd4-4807-9d47-24f40d1aadb2 · outbound

This paper cites Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024.

Physical Informed Driving World Model Moonshot: To- wards controllable video generation and editing with multi- modal conditions, 2024

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.783763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.463513Z digest=sha256:a5172089ef8c5cb2829279e36e4734274714fb1ddfcd6b7def50a8afa5295487

Observation 85472ba3-cd49-4454-86d6-6f450f65df25 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Physical Informed Driving World Model Adding conditional control to text-to-image diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.467551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.467551Z digest=sha256:5f99b521d1eefc577dd6b20cc18ac7387fc6ccc44d0caa31635c642e6bdfeb4b

Observation d760a2ca-8a8a-463e-999a-4977e19d6fc3 · outbound

This paper cites Collaborative and adversarial network for unsupervised do- main adaptation.

Physical Informed Driving World Model Collaborative and adversarial network for unsupervised do- main adaptation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.764198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.471434Z digest=sha256:36c612ef28aa62d71454acc896d15bf5fa18b2aaa0a5535b735825adc5840de4

Observation 4b2863e2-328f-4295-9fad-e79343f57c3b · outbound

This paper cites DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation.

Physical Informed Driving World Model DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.475852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.475852Z digest=sha256:7fdb786514d63d295ff75ac9d29c868c005cf4e2400066206a2e2d8657a1b0e2

Observation 3768ceed-6242-4fd5-851f-519b592fb326 · outbound

This paper cites GenAD: Generative End-to-End Autonomous Driving.

Physical Informed Driving World Model GenAD: Generative End-to-End Autonomous Driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T17:58:58.480102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:58:58.480102Z digest=sha256:4ce113d0d5fe620aa0f71ae3b286906e7b85dc8575cb337fc2dd86b8f3a76137

Observation bd638ff6-3a72-44aa-a7ac-bd090241bf8c · outbound

This paper cites Open-sora: Democratizing efficient video production for all, 2024.

Physical Informed Driving World Model Open-sora: Democratizing efficient video production for all, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.752042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.484128Z digest=sha256:3853462e2db551d8a2676379a0baaa97361184bbfca7692bb9d5ba94d6feb57a

Observation 16d0aac3-1ead-4981-b815-6823e7a80c92 · outbound

This paper cites patchified.

Physical Informed Driving World Model patchified

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.738519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.488155Z digest=sha256:005c70b4fabf32f2a63f478216598a9f7b66c79e37e71d3b1fb8f360c8f14102

Observation 35f79062-b189-47a1-960b-f48e0bd19b09 · outbound

This paper cites Simulation-to-Real Visual Translation.

Physical Informed Driving World Model Simulation-to-Real Visual Translation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:58:58.724458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:58:58.492878Z digest=sha256:8a497e079abe9def6e764bf7e0f3c33d5e182b8d83b125db29786f13c00aece7

Pith citing papers

Observation 41ebb73c-9c33-49a7-98e5-dd8f28167298 · inbound

A Survey of World Models for Autonomous Driving cites this paper.

A Survey of World Models for Autonomous Driving Physical Informed Driving World Model

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-10T18:31:52.984933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:31:52.984933Z digest=sha256:aefbf9c8a5838c9198e571a7cdef8d30750e070392433f09c4bb9e1a0589a578

Observation 3afa36f0-f4de-441e-aba1-ea4e330d1edf · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment Physical Informed Driving World Model

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.799714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:788091bb85042c0d805b865e9adeb7d71dec0bf1497fd317195f616fa08768b9

Observation 24f0798c-92af-4e58-bc33-13b694187851 · inbound

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities cites this paper.

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities Physical Informed Driving World Model

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-04T20:52:06.587203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:52:06.587203Z digest=sha256:278713961c6c488e63fb39e957f8a1fa02a3fedb4d31450cae5f70cc202ecdfe

Observation 2c5cbbd6-3523-43cc-a13a-25dbb91d6b98 · inbound

A Comprehensive Survey on World Models for Embodied AI cites this paper.

A Comprehensive Survey on World Models for Embodied AI Physical Informed Driving World Model

Reference 214

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:54.374022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:12:54.374022Z digest=sha256:d25fdf6f90e950ab9a6d28f66e468c8cee3c9175d1924e235b9948545a3f8c56

Observation a03e8934-99c5-43f2-bf1e-d4066f579245 · inbound

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World cites this paper.

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World Physical Informed Driving World Model

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-03T17:02:42.473873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:02:42.473873Z digest=sha256:92313933f906b5df9520869a920a49415f87010262a8e991b3b1ecbb232d3c8d

Observation dabfd1d9-faf5-4c98-aa6b-3507fe97cfca · inbound

MultiWorld: Scalable Multi-Agent Multi-View Video World Models cites this paper.

MultiWorld: Scalable Multi-Agent Multi-View Video World Models Physical Informed Driving World Model

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:09:08.093011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T05:06:11.514186Z digest=sha256:e0542af7ff7d54a9723842a6368950b96029baee8f9cfa37446e897a134448f1