Pith. sign in

Paper Citation Record · LEDGER

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

As of 5 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 6 inbound Pith citation observations for arXiv:2512.23180.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.23180 v3

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T17:06:34.398973Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T13:31:10.576156Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T17:08:00.948705Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact23
  • verified fuzzy40
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 41be73b1-d803-4c8f-9772-7e0dc7f0f13d · outbound

This paper cites GPT-4 Technical Report.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation GPT-4 Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:10:25.121871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:0ba0c1cba10939a0fab0c74972620e67c0faf13662ce3563635a31acf77fb28a

Observation c3e9ddd1-9cf0-4fff-927c-4772dfa87640 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:10:25.095962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:8504d05f8c3cf382d90fa87032de366b93268a56f6599b20e7951a70d4e8fe71

Observation 5284ef9e-cd56-4670-8711-e36663f8e5d3 · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation nuscenes: A multi- modal dataset for autonomous driving

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.072050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:68bb371ad2582a4ce3e1e6d908ed79e1ea5b83b174336bb2b8c1b1a5eda99e37

Observation 2505e440-b2ce-441b-a6fb-784fb1ed269a · outbound

This paper cites Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Periodic Vibration Gaussian: Dynamic Urban Scene Reconstruction and Real-time Rendering

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.109113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:76b5638a791e18ef967bce9e7b748a9d551511dbc00a74ce0a0f0249c9cff693

Observation c058bf0a-32c7-4f3f-8a6a-ec472da5f164 · outbound

This paper cites SN-LiDAR: Semantic Neural Fields for Novel Space-time View LiDAR Synthesis.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation SN-LiDAR: Semantic Neural Fields for Novel Space-time View LiDAR Synthesis

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.131282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:eff2fdd831c8d0ea8dcb4c84091a2e345f2435a9b944a3c2fadf11b015be8bd3

Observation 7b997547-040e-4f6f-8ae8-bb43eeb053d9 · outbound

This paper cites How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.Science China Information Sciences, 67(12):220101.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.Science China Information Sciences, 67(12):220101

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.956530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:97b706f8fadc414ec36fd673d0b4dd0520090adee153d775e575b006eb06881a

Observation 53a8afd7-33b1-4cff-98aa-c60580fdcf2c · outbound

This paper cites Omnire: Omni urban scene reconstruction.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Omnire: Omni urban scene reconstruction

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.953122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:8fcfabf75b9729b9b80755aa2940ba277c7e6440674ca4cbd688db7424bc6363

Observation 4cd816e4-7b20-48c0-ac77-4bc1a69215b4 · outbound

This paper cites ProSGNeRF: Progressive Dynamic Neural Scene Graph with Frequency Modulated Foundation Model in Urban Scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation ProSGNeRF: Progressive Dynamic Neural Scene Graph with Frequency Modulated Foundation Model in Urban Scenes

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-13T02:17:00.989019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:b96c8fe44cd8f059733a16f35712ae1bdd68f729c2009fce03e776541c0abac2

Observation 917f9d43-2d73-4b9a-b2e5-2bffed6dd604 · outbound

This paper cites Plgslam: Progressive neural scene represenation with local to global bundle adjustment.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Plgslam: Progressive neural scene represenation with local to global bundle adjustment

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.959863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:ceb7ad2f4941b8544c20e76c882faf11e12a5a827a7e8e67b636d816738e20a4

Observation bc5b648f-dfdd-44ad-a3a4-db41b2c98c47 · outbound

This paper cites What is the best 3d scene representation for robotics? from geometric to foundation models.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation What is the best 3d scene representation for robotics? from geometric to foundation models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.100430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:782200f155ffe2e4dfb9afe45dd92a7528f3c9162b4d2c5bd37a6ea361cfb9cb

Observation 885d050a-db71-4834-b741-95ff520237b4 · outbound

This paper cites MCN-SLAM: Multi-Agent Collaborative Neural SLAM with Hybrid Implicit Neural Scene Representation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation MCN-SLAM: Multi-Agent Collaborative Neural SLAM with Hybrid Implicit Neural Scene Representation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.136388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:b6565378e0ad806fac9627b82d3dce1f588be17552f3a29c13ffe13f2a29eb6a

Observation 781b0892-0ffd-437d-957f-cc935670672f · outbound

This paper cites Mne-slam: Multi-agent neural slam for mobile robots.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Mne-slam: Multi-agent neural slam for mobile robots

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.950372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:85933d1f39e27f93ffada14a97e93dc3c65a01984ab96fe8d52453dad08f2b82

Observation 0b48d5e7-a59c-467c-91b8-2b388745e3e6 · outbound

This paper cites an unresolved cited work.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-21T17:10:26.068813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:c330bc4cc1bc98b8f963b294d69bb12822e50b965e29b4e381240410223f6a34

Observation 624ad7d5-8587-4c24-8d90-84b11c97f0a9 · outbound

This paper cites Vpgs-slam: V oxel-based progressive 3d gaussian slam in large-scale scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Vpgs-slam: V oxel-based progressive 3d gaussian slam in large-scale scenes

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.192004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:d9e1a2b85f2af385e1a74efd38f6f5828235c7b06a3386f2bcd096446808f60e

Observation af20115f-79e0-430d-a306-b26fc0222edb · outbound

This paper cites Scaling recti- fied flow transformers for high-resolution image synthesis.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Scaling recti- fied flow transformers for high-resolution image synthesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.065974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:9a337458e2f6fc878c0433b8c4f34129bd4be16506494d30c666ccb319b6cd1b

Observation 5d4f4b16-3e17-41b0-a456-f9dc19645bc1 · outbound

This paper cites MagicDrive: Street View Generation with Diverse 3D Geometry Control.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation MagicDrive: Street View Generation with Diverse 3D Geometry Control

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.187814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:8c844feb01a5abc08ee349254fb2f090c0dff0fb40ab711731c89175cfc58694

Observation b4d6c827-e202-40dd-b3f3-189e59bac8ae · outbound

This paper cites Vista: A generalizable driving world model with high fidelity and versatile controllability.Advances in Neural Information Processing Systems, 37:91560–91596.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Vista: A generalizable driving world model with high fidelity and versatile controllability.Advances in Neural Information Processing Systems, 37:91560–91596

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.062079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:0fb934db5f30d589ff9d7c27393f757bd3a9514e289a28b1f74b58d0c5792228

Observation 15da5d56-1ed3-45fb-8f84-1e11931f69be · outbound

This paper cites Mocount: Motion-based repetitive ac- tion counting.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Mocount: Motion-based repetitive ac- tion counting

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.058765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:82f2ef52978f8f42a54ec8a099291f191a87bcf8d85097c7493a2fe4169b8307

Observation 8d4dfb09-b038-4e2b-b586-5b735883cde3 · outbound

This paper cites Dist-4d: Disentangled spa- tiotemporal diffusion with metric depth for 4d driving scene generation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Dist-4d: Disentangled spa- tiotemporal diffusion with metric depth for 4d driving scene generation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.055578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:d59b546eeeb8a8b7aeb8df08b18ebdbbeda0d31ecb017ec382d86ab349476988

Observation 6dbf3dd0-572f-40e1-8dcf-9a7bb3181a74 · outbound

This paper cites DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.148944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:abd28a5245d7a2933c8701cd32e4484315fc4719f430aa034deba9419852acd9

Observation 86b7e462-9ced-4beb-bd69-2583450238f1 · outbound

This paper cites GaussianVLM: Scene-centric 3D Vision-Language Models using Language-aligned Gaussian Splats for Embodied Reasoning and Beyond.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation GaussianVLM: Scene-centric 3D Vision-Language Models using Language-aligned Gaussian Splats for Embodied Reasoning and Beyond

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.178774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:aba5763478e7a912946f7a4558edb222d42b4377c3d6fa6e35bd4fe01b3c83dd

Observation 912388c2-61f7-4054-bad7-55fe772b783d · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.Advances 17 in Neural Information Processing Systems, 36:20482–20494.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation 3d-llm: In- jecting the 3d world into large language models.Advances 17 in Neural Information Processing Systems, 36:20482–20494

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.052326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:e5e884aaba48e382f8bb541bc28053758a19e59d3e71fd5316e606806808c857

Observation 58fdc6c1-aacf-42d3-8dd1-ac838e2a0feb · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation GAIA-1: A Generative World Model for Autonomous Driving

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:10:25.153507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:8b212049fd3172e74f05a293e0c8303160e125aaa808eecc57087d369a0910cd

Observation d92bbf53-ac26-4ddd-ab65-9a832bd97808 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Trans.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation 3d gaussian splatting for real-time radiance field rendering.ACM Trans

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.049436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:bb270856ff0ea6212283dbfc66536a986a3d48cecd58a9364cd8550c66c6afa7

Observation 5724e774-4ef0-4ad0-b01a-c8b1da45ae31 · outbound

This paper cites Segment any- thing.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Segment any- thing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.046450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:7fb4b71130f2500109ca4cd3dec8c37667bb9f74a04f02ca6911a42adb0d41e9

Observation 2d6c893e-77f5-41c8-9696-3fb7812e0860 · outbound

This paper cites 3D and 4D World Modeling: A Survey.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation 3D and 4D World Modeling: A Survey

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-21T03:22:29.733321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:ea26e81273e7dec8e105bc0f2c6c2fb3c51227ce1edd58a979e249710d6728aa

Observation 77071d20-82d4-42c7-bf1a-3120ab6ae60c · outbound

This paper cites Uniscene: Unified occupancy-centric driving scene generation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Uniscene: Unified occupancy-centric driving scene generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.043249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:3d910ce2c5dd6791fb42244465ae70c583d4dc3d8b9e0461251833609f7f596f

Observation d5dd6ddc-ea0d-4dd4-9e53-5b3820216edb · outbound

This paper cites Human motion instruction tuning.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Human motion instruction tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.040268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:2c474d0dc58e8b30e462910af82a38bfabf3201b8691dc385432a1cacffa6ca3

Observation 0fb91e00-d989-4072-98b4-440e7d74e00b · outbound

This paper cites Bevformer: learning bird’s-eye-view representation from lidar-camera via spatiotemporal transformers.IEEE Transactions on Pat- tern Analysis and Machine Intelligence.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Bevformer: learning bird’s-eye-view representation from lidar-camera via spatiotemporal transformers.IEEE Transactions on Pat- tern Analysis and Machine Intelligence

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.037040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:e6fc7b639dfdf6e7b4ae7bf4cfe3de178c18225c5f5fa316ea7ad112aa5c93f4

Observation 65e38c08-4e44-44bd-96a7-3768949905e8 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Rouge: A package for automatic evaluation of summaries

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.034077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:7bc55f19ed2e9b190f9e4d44140ed79f78fdda8ccd2714f8f85e1e3bca753c73

Observation a2b336cb-d676-4884-951b-110e4f089cda · outbound

This paper cites Improved baselines with visual instruction tuning.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Improved baselines with visual instruction tuning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.031555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:f66a588c40e83bc8163e45a76a5612ad5d1d24e28e100ae8573e49d57b1746b2

Observation 77d3fde9-1b54-4a03-b293-00e2127e77ed · outbound

This paper cites Petr: Position embedding transformation for multi-view 3d object detection.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Petr: Position embedding transformation for multi-view 3d object detection

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.028720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:c81b31aba00e51a22e0ea45dec46d54af2573dfac19da777a3e76c2b72d86886

Observation a9a962a8-b05d-43cf-8abb-1bcb2e7dce43 · outbound

This paper cites Dreamdrive: Generative 4d scene modeling from street view images.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Dreamdrive: Generative 4d scene modeling from street view images

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.025930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:eb960f73ca72888ae0fa7162e1892814a6c353a66728f7b1e156cfe341c3bf43

Observation 972ac40c-f240-4684-a435-3610640f5757 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.Communications of the ACM, 65(1):99–106.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Nerf: Representing scenes as neural radiance fields for view syn- thesis.Communications of the ACM, 65(1):99–106

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.022990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:3e597d068c4207a41237a6a23eccb5a1a11c0026a4855b8527628264ba7c2dad

Observation cd8d2633-665f-4da9-9a18-65f516259708 · outbound

This paper cites Neural scene graphs for dynamic scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Neural scene graphs for dynamic scenes

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.019645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:6f366ef7852304b4f7283fd9398a1a52595a5976be33f2cb464572e7a1efa1e0

Observation 204d4794-79fe-49ee-9cf3-26ce48aaea63 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Bleu: a method for automatic evaluation of machine translation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.016813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:e773aec71ef018c54e903628cd9e2b575b0bf5950800842c027e8ad61c8192f7

Observation a5ecf34f-81cc-4998-ac08-a0498dcbcc04 · outbound

This paper cites A lesson in splats: Teacher-guided diffusion for 3d gaussian splats generation with 2d supervision.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation A lesson in splats: Teacher-guided diffusion for 3d gaussian splats generation with 2d supervision

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.013952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:fe3a028684ff0e62423e313beab1b4f6598f1b514af338cbae640f1fbc340087

Observation ec19c342-9138-4638-ae7d-4f5dff1f7e68 · outbound

This paper cites Desire-gs: 4d street gaussians for static-dynamic decomposition and surface reconstruction for urban driving scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Desire-gs: 4d street gaussians for static-dynamic decomposition and surface reconstruction for urban driving scenes

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.010944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:8f5262a600812561160a41c4bf4599eaab0274b00951201eab4d0e0aab8cc358

Observation 43f3beeb-026e-4386-b4f3-88ac326ee877 · outbound

This paper cites Langsplat: 3d language gaussian splatting.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Langsplat: 3d language gaussian splatting

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.007857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:084ac472f53fcd771ad5569998db6db3fa29f54fd0c5b08299dd85e8d6628dbe

Observation 92c859f4-c4ca-4391-b504-242bf9d40226 · outbound

This paper cites Drivelm: Driving with graph visual question answering.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Drivelm: Driving with graph visual question answering

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.005001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:42ba395b66cc447a60f3a6be38afdb8743da4fd52dc895f2a13351706871cf21

Observation 26ebf199-43b9-4ab6-97de-96e7ec64d2ea · outbound

This paper cites Driv- ingforward: Feed-forward 3d gaussian splatting for driving scene reconstruction from flexible surround-view input.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Driv- ingforward: Feed-forward 3d gaussian splatting for driving scene reconstruction from flexible surround-view input

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:26.001855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:d42235f94566ca0b401929f58731b6020946821257e008c4447400a7e3b9a03d

Observation 89696230-600f-455b-a254-00e1b1ae629f · outbound

This paper cites Suds: Scalable urban dynamic scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Suds: Scalable urban dynamic scenes

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.998893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:451fca0db38aff0f3c2ef9cd92faf75d0719cb79356cd311a46b994c285aedb6

Observation 30cf4b65-afde-4d84-8480-50f97039f731 · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Cider: Consensus-based image description evalua- tion

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.995959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:eff21633f276bd60031d18368a08b27caaef845374ace4a893c8322435c8578b

Observation d6d7c433-764e-4e6e-85e6-d8585c2917e0 · outbound

This paper cites Freevs: Generative view synthesis on free driv- ing trajectory.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Freevs: Generative view synthesis on free driv- ing trajectory

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.993538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:6693b944b930217180bfb6160dff1a092db3a9cf7e92f9cec73fd397e08bcd5a

Observation 4bad99bb-55ef-4021-8ee7-0dc5ca1df234 · outbound

This paper cites Omnidrive: A holistic vision-language dataset for au- tonomous driving with counterfactual reasoning.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Omnidrive: A holistic vision-language dataset for au- tonomous driving with counterfactual reasoning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.990708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:14be8f35ee0d7d56e43c85de863999b8a34862ed13c53b222bb39994cee3a1b8

Observation 570ef81b-118f-4e0b-9621-5568a715d339 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.987613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:4e34a2465c7b17a1eb9b2a2cf593b521ddeb6a6731f4942452487fb528ce72eb

Observation e2751359-d733-4f8b-8c42-6fffe58c58fd · outbound

This paper cites Learning to Tune Like an Expert: Interpretable and Scene-Aware Navigation via MLLM Reasoning and CVAE-Based Adaptation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Learning to Tune Like an Expert: Interpretable and Scene-Aware Navigation via MLLM Reasoning and CVAE-Based Adaptation

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.117502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:59d7d34aa9e0c3549d177760fffdc00f8e6414cba4a8e435a83d7b7ac024f68f

Observation 0c78d41a-6594-41db-8417-259c653bfa7d · outbound

This paper cites FreeDriveRF: Monocular RGB Dynamic NeRF without Poses for Autonomous Driving via Point-Level Dynamic-Static Decoupling.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation FreeDriveRF: Monocular RGB Dynamic NeRF without Poses for Autonomous Driving via Point-Level Dynamic-Static Decoupling

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.113160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:f5c7aa73a339fcb27e1499caf13b7796bcc1358f083b26f1b1fc043a6b60033a

Observation 4f9e0473-1be6-4a87-b4c6-c8487cfaf7d7 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.984189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:c34e6a745cda57e3fd636bff352e864c70bb970eec4b440396bbfb35894acf40

Observation d5a040d2-0be0-445a-ad58-c6785259571a · outbound

This paper cites Cape: Camera view position embedding for multi-view 3d object detection.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Cape: Camera view position embedding for multi-view 3d object detection

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.981152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:59af236eb23b3be7742ce364ceac2d89597d994737ea442869f077950ce0a028

Observation ac527803-74d2-47a8-ba17-3021452ec1e4 · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Let- ters.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Let- ters

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.978254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:516d50bd6c0e717e67af847f283730dd805f5c048d26230550912e7954d39505

Observation f4a4d8d2-6abd-45f2-a018-5bf6acd2293a · outbound

This paper cites Drivingsphere: Building a high-fidelity 4d world for closed- loop simulation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Drivingsphere: Building a high-fidelity 4d world for closed- loop simulation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.974931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:b45f95f6c80ae6f87552798cae6eaa71d494930402eae2fb4f220bcd720c0213

Observation b7edaaa9-10b7-40d2-8efd-cab150dfa4b5 · outbound

This paper cites Street gaussians: Modeling dynamic urban scenes with gaussian splatting.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Street gaussians: Modeling dynamic urban scenes with gaussian splatting

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.971991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:5a8cbec83403ba3c73504adc8c0376e3ba7604f2b6b0508085cb9bf58c5748be

Observation d43fb6d7-b8a2-4d71-a1e9-0a91d583166c · outbound

This paper cites Qwen3 Technical Report.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Qwen3 Technical Report

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:10:25.144067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:e460332f19c51bb7c92206eb99ef23023c6dcc00fa7efba1502c25bc158cee57

Observation c6e30895-495d-4f34-a1c1-f4e46c8085d4 · outbound

This paper cites EmerNeRF: Emergent Spatial-Temporal Scene Decomposition via Self-Supervision.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation EmerNeRF: Emergent Spatial-Temporal Scene Decomposition via Self-Supervision

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.104651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:1a6fd8810bef7b60a6d1d683d9a5c5379c67cf17c272b3f1c7ec2e7caee531de

Observation 4307bcd8-2281-4233-92a0-41ccddbab1e1 · outbound

This paper cites STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.126719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:07e34795ccf0bc98c8a8197c644a0430e9b3370fc5bc4f8ad1589f2aae0aa142

Observation 4892945c-1197-4045-b17e-0b6982544b6f · outbound

This paper cites Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene Reconstruction.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Deformable 3D Gaussians for High-Fidelity Monocular Dynamic Scene Reconstruction

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.183203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:c9b4d97a03e8e1058c3bedcd4e714ca9f0898ece3b5a6ef5728d594514831b45

Observation abc6635c-121c-42e2-b1f7-518a0714f065 · outbound

This paper cites Visual point cloud forecasting enables scalable autonomous driving.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Visual point cloud forecasting enables scalable autonomous driving

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.969213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:2f9815d8c67390f540ea44d369188df226581c103633187d27637c81c9eb4c6a

Observation 1448b9dd-b4fa-4bac-a880-84ad3ac2ee60 · outbound

This paper cites Drivedreamer-2: Llm-enhanced world models for diverse driving video generation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Drivedreamer-2: Llm-enhanced world models for diverse driving video generation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.966280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:319124975a915be6e322fdbd27ad8184efc517966262e722fa3083814d9da069

Observation a561ba29-651a-49e2-bdc8-6f643e12bc5e · outbound

This paper cites Extending Large Vision-Language Model for Diverse Interactive Tasks in Autonomous Driving.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Extending Large Vision-Language Model for Diverse Interactive Tasks in Autonomous Driving

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.162426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:cdcc39662e8e938b106cae2c2773fe6fa0ddb2cf40ef610a6ee0b4b24f98d334

Observation c0db86ef-5a47-47d0-8270-c69c21a01831 · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:10:25.157527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:54cdcbde07e1ef601e72657ba04d215de63d921a4a53484a53c2b31c43ac906d

Observation df848589-0fbf-4912-8a6c-53760c492a0d · outbound

This paper cites Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation Drivinggaussian: Composite gaussian splatting for surrounding dynamic au- tonomous driving scenes

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:10:25.963205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:e15c9e34f2fd6ce67844c81340065f197266b9ad45e1b72348721df08f838b9e

Observation 76b4ae62-c7e5-42df-959a-a0c1abf1c81a · outbound

This paper cites HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.174828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:b1131de4c1087e9120cbd5530ac2e9a557fa2f93d69a5d0e081d71757a2a5238

Observation 916059ab-6930-4e5e-bc2a-be41c95284dd · outbound

This paper cites MuDG: Taming Multi-modal Diffusion with Gaussian Splatting for Urban Scene Reconstruction.

GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation MuDG: Taming Multi-modal Diffusion with Gaussian Splatting for Urban Scene Reconstruction

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:10:25.170652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T17:06:34.398973Z digest=sha256:ed1945125a3aa58fca9558f6a2b6af6ca1055edadb9325084ed45f0a112bebd2

Pith citing papers

Observation ea80281b-b8c8-40b7-a35f-c1b13e144ed7 · inbound

Evolutionary Physics-Informed Temporal Fusion for Lane-Change Intention Prediction cites this paper.

Evolutionary Physics-Informed Temporal Fusion for Lane-Change Intention Prediction GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T13:31:10.576156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:31:10.576156Z digest=sha256:de3786859c55720ae4bac79e71bb06933234b2ed855c67c6523bd2054815ab8b

Observation 9b3e77ca-d219-479a-a9ed-f9b9dfc89605 · inbound

Spatial navigation in preclinical Alzheimer's disease: A review cites this paper.

Spatial navigation in preclinical Alzheimer's disease: A review GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T19:53:53.219607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:53:53.219607Z digest=sha256:92f2987a72df31ccb0186d45cd8d5300313c5aa8f25d12c85185e92923e95608

Observation 5472f68e-b5be-45cc-862f-2f8792081ba3 · inbound

DINO-VO: Learning Where to Focus for Enhanced State Estimation cites this paper.

DINO-VO: Learning Where to Focus for Enhanced State Estimation GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:03:01.148630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T17:05:52.493660Z digest=sha256:ca481e176143e4eca7a640389777a8cd38490f22dc73c9cda1717f4ff4983ab6

Observation 5e55bb5d-c3d0-4818-825d-0134a35f8e41 · inbound

Appearance Decomposition Gaussian Splatting for Multi-Traversal Reconstruction cites this paper.

Appearance Decomposition Gaussian Splatting for Multi-Traversal Reconstruction GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:03:01.148630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:11:56.114485Z digest=sha256:d4e9af0242fa301ec98fc957be51ef359c856dfc94346e502c8c3a62a3f1f658

Observation 613b953d-2777-404a-9919-556c9747648b · inbound

LLM-Augmented Traffic Signal Control with LSTM-Based Traffic State Prediction and Safety-Constrained Decision Support cites this paper.

LLM-Augmented Traffic Signal Control with LSTM-Based Traffic State Prediction and Safety-Constrained Decision Support GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:03:01.148630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T05:59:22.177960Z digest=sha256:8049747fcb01375837a83830f5a4dc2b3f89b1b9be7a0eb7c1cf015556fa5aa0

Observation 8256753b-ba6b-4e00-92b4-ce3ec79adf70 · inbound

CoSAG: Compact Semantic Anchor Gaussians via Training-Free Rate-Distortion Coding cites this paper.

CoSAG: Compact Semantic Anchor Gaussians via Training-Free Rate-Distortion Coding GaussianDWM: 3D Gaussian Driving World Model for Unified Scene Understanding and Multi-Modal Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-14T13:19:08.066069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T13:19:08.066069Z digest=sha256:8d1cf9b76b3871e1cd2743bd36488776cd8612267b4195b74c27df5435e577ad