Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:53.713132Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 12 inbound Pith citation observations for arXiv:2506.07497.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:53.713132Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T08:57:08.231458Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T19:28:52.737750Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a48e3d34-be12-4c25-8957-eee046ed32f6 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Qwen Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cf9c107-4208-4985-8261-389c8dbe4bda · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Transfusion: Robust lidar-camera fusion for 3d object detection with transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 951a427e-9b2f-4954-aef0-924c0dafc292 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Camera-lidar inte- gration: Probabilistic sensor fusion for semantic mapping.IEEE Transactions on Intelligent Transportation Systems, 23(7):7637–7652, 2021
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4a22c19d-e3b2-4e1d-825a-02fd06bffdac · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency nuscenes: A multimodal dataset for autonomous driving
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda1960b-60af-4262-b5bc-26c8006656af · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a21c1d5d-5742-4009-8023-4dd70a137bf5 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Scaling rectified flow trans- formers for high-resolution image synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10daedaf-dc4b-4d75-8b65-b0d607bf7e60 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Trafficgen: Learning to generate diverse and realistic traffic scenarios
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation da919632-7ffd-434f-b87f-79a35b2122bb · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2be098a-fc3e-4f7b-be2c-1cabcc112221 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e36da3c9-0cb7-4ca7-adad-16d9513f481a · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency MagicDrive: Street View Generation with Diverse 3D Geometry Control
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 973f4d24-146c-4556-84f1-cc72adaa1088 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2f7d725-15d0-43c3-b676-7d7fdc604d78 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Vision meets robotics: The kitti dataset.The international journal of robotics research, 32(11):1231–1237, 2013
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8de2b146-c2f5-4380-bf0a-8604a0b1059a · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b6be0d0-9fd9-4173-ba29-d76ec2c4ef50 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 913b628b-250c-4159-8935-e1ae5e0e3e78 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency CoGen: 3D Consistent Video Generation via Adaptive Conditioning for Autonomous Driving
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02560e80-bbaa-447a-bb00-5ed5cdf6d0cd · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency DiVE: DiT-based Video Generation with Enhanced Control
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e9d4c05-0c91-43f8-8ef7-9336d47a01dd · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Point cloud forecasting as a proxy for 4d occupancy forecasting
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2d47840-adea-440e-be8f-8bc4740ac7c2 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency UniScene: Unified Occupancy-centric Driving Scene Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ae218aa-f448-430e-afc0-e5b031fe2903 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Deepfusion: Lidar-camera deep fusion for multi-modal 3d object detection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 97817166-7d39-4b11-a189-dc96bde9cd80 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Bevfusion: A simple and robust lidar-camera fusion framework
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd90cf4-5e54-4f8d-b008-e13244d031a6 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9737325c-0413-4e95-b419-23730d09000f · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Swin transformer: Hierarchical vision transformer using shifted windows
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c857f0af-00ae-45cb-baf7-11f143b1bac9 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Wovogen: World volume-aware diffusion for controllable multi-camera driving scene generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2915b128-9a06-49ca-a5d1-64de3241a0e5 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f7dda2d-9abb-4652-9afd-a3f97aba60be · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Nerf: Representing scenes as neural radiance fields for view synthesis
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ceec62-ab6d-467b-bedf-2291c9799b41 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 64bdebe8-b95b-4f1d-ae90-fe2176caec3c · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Scenario diffusion: Controllable driving scenario generation with diffusion.Advances in Neural Information Processing Systems, 36:68873–68894, 2023
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 22009dd1-b790-4965-b85b-7a378d5fc560 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Pointnet: Deep learning on point sets for 3d classification and segmentation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cbb251d-e85a-437c-8743-6dbdd8b3547e · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Towards realistic scene generation with lidar diffusion models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a17f0717-87f3-446c-af2d-92414bfc62b4 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Scalability in perception for autonomous driving: Waymo open dataset
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 875c1f86-1877-4056-8221-4cb9cd0c7973 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Drivescenegen: Generating diverse and realistic driving scenarios from scratch
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2edd54ca-b8e8-4128-b6af-f50d2f8d4774 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Street-view image generation from a bird’s-eye view layout.IEEE Robotics and Automation Letters, 2024
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9391a233-0bf3-48f2-a4ce-b5a7b6ebffa3 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Fvd: A new metric for video generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf12d541-a77e-45d6-b49d-d0f521364c61 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b1ba8b9-7c0f-4cd7-a855-94eeba645fe7 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Drive- dreamer: Towards real-world-drive world models for autonomous driving
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8225b22e-44d9-4c4d-8b80-d6b4a97f9960 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Driving into the future: Multiview visual forecasting and planning with world model for autonomous driving
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9488ad-af98-47a8-9014-7efb31e322e6 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Lidenerf: Neural radiance field reconstruction with depth prior provided by lidar point cloud.ISPRS Journal of Photogrammetry and Remote Sensing, 208:296–307, 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2746e56b-f07c-4aa8-80ce-5d50bf531a20 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Panacea: Panoramic and controllable video generation for autonomous driving
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7c8a747-501a-4309-9bef-d5326deff41e · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 041e5865-e3c7-46e1-b6ff-e3b6f39493a9 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Driver lane change intention recognition based on attention enhanced residual-mbi-lstm network.IEEE Access, 10:58050– 58061, 2022
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0b07300e-7184-40d3-8449-2acb762681fe · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Point-nerf: Point-based neural radiance fields
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f18b92f-e605-4921-9927-d8ac85237541 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1639931f-30c7-4fe4-aa52-8d6268b93b7e · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Visual point cloud forecasting enables scalable autonomous driving
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f38f51f-333a-43ae-b05d-84c017338b22 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99a95cf3-3760-4543-bbcf-64ba67862d16 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 469a0893-7342-48a4-803a-2ce7a1c2b7b0 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Drivedreamer-2: Llm-enhanced world models for diverse driving video gen- eration
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aa96d767-5e76-4214-afd4-77db0d641785 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 580001f9-4d2b-467d-aaa4-6e5c54160ed1 · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency V oxelnet: End-to-end learning for point cloud based 3d object detection
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4964914-1372-4e3f-a0b4-e9c6788d3f3b · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Lidardm: Generative lidar simulation in a generated world.arXiv preprint arXiv:2404.02903, 2024
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 883a735b-0381-49e1-9e5b-1cb44fe15c5b · outbound
Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency Daytime",
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1f4e4ccf-3a5d-48c7-bd29-26a10afca9d9 · inbound
OmniNWM: Omniscient Driving Navigation World Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc0aecb1-ed81-4793-ae62-f216b225084a · inbound
A Survey on the Applications of Generative Artificial Intelligence in Automated Driving Systems Test Scenario Generation Methods Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37eb6db4-a748-4454-8e71-c49ba98fabb7 · inbound
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 41783d20-7bf4-4541-b45e-20fb32be0b90 · inbound
UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6510c82b-a3cd-4d81-946b-6fd8ced5044a · inbound
From Seeing to Simulating: Generative High-Fidelity Simulation with Digital Cousins for Generalizable Robot Learning and Evaluation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ca8be82c-36cf-427f-82a8-afccb44d388f · inbound
CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cabd6262-800a-4b4f-b89c-cab10128566b · inbound
CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3086f0b7-76e7-4d65-9840-fb635026f0cd · inbound
Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 25f7670b-e538-47fb-8b7c-c1bd06464768 · inbound
Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0bfa4449-05bd-4707-a3ea-fe2515e5471e · inbound
OmniDrive: An LLM-Choreographed Multi-Agent World Model with Unified Latent Co-Compression for Multi-View Driving Video Generation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 89427131-e1a6-4272-901f-3792c54795eb · inbound
ReWorld: Learning Better Representations for World Action Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dceccfac-7cfb-4f7f-9062-8aa62f9a58a0 · inbound
UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.