Pith. sign in

Paper Citation Record · LEDGER

Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2506.07497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07497 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:57:08.231458Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T19:28:52.737750Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1f4e4ccf-3a5d-48c7-bd29-26a10afca9d9 · inbound

OmniNWM: Omniscient Driving Navigation World Models cites this paper.

OmniNWM: Omniscient Driving Navigation World Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T08:57:08.231458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:57:08.231458Z digest=sha256:ddc0e32330ea45f524e398867a98770cbeee2d074b7c5dba296617dae70051db

Observation bc0aecb1-ed81-4793-ae62-f216b225084a · inbound

A Survey on the Applications of Generative Artificial Intelligence in Automated Driving Systems Test Scenario Generation Methods cites this paper.

A Survey on the Applications of Generative Artificial Intelligence in Automated Driving Systems Test Scenario Generation Methods Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T15:54:39.583328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:54:39.583328Z digest=sha256:01f276a5f1c57b1a77918a22dd437055de7d5bd471a4d87e05433d1b10591072

Observation 37eb6db4-a748-4454-8e71-c49ba98fabb7 · inbound

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World cites this paper.

DriveLaW:Unifying Planning and Video Generation in a Latent Driving World Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:38:21.112262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:34:39.518649Z digest=sha256:20855af8b68d81b61ad31848bfe5346c1a61e4c08aefae93c9d9f630f24174fc

Observation 41783d20-7bf4-4541-b45e-20fb32be0b90 · inbound

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models cites this paper.

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T20:36:09.847931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:36:09.847931Z digest=sha256:0f70191de24bba8d20361ed06778eb2241c414812004a6e411ac170b59f1925e

Observation 6510c82b-a3cd-4d81-946b-6fd8ced5044a · inbound

From Seeing to Simulating: Generative High-Fidelity Simulation with Digital Cousins for Generalizable Robot Learning and Evaluation cites this paper.

From Seeing to Simulating: Generative High-Fidelity Simulation with Digital Cousins for Generalizable Robot Learning and Evaluation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:48:01.778224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T08:45:46.944379Z digest=sha256:82e57d5285ce8ed99085ec91274b9c5019f6dc42da2ecc20bf08bf41905c8dfa

Observation ca8be82c-36cf-427f-82a8-afccb44d388f · inbound

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving cites this paper.

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:56:28.605367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T03:48:36.717026Z digest=sha256:c61c632e47c80f7485a1e030b0e25e264001f26508666bf9d549188d9af8849c

Observation cabd6262-800a-4b4f-b89c-cab10128566b · inbound

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving cites this paper.

CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:38:00.560018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:36:52.396245Z digest=sha256:fcb48aa99e432f4ff1170020da67265cdbff6b167cc418d50e30a3ece30d171c

Observation 3086f0b7-76e7-4d65-9840-fb635026f0cd · inbound

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving cites this paper.

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:01:09.043469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T05:56:45.641080Z digest=sha256:9b7b6370db041713e4db766511ec6745df76ce7c0cab0c3d0142bfe8858048a2

Observation 25f7670b-e538-47fb-8b7c-c1bd06464768 · inbound

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving cites this paper.

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:44:55.995325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T16:40:24.858861Z digest=sha256:891dcb74ee1c47232e3369f1cbaf737487552410b879ee2c89c323274d19cce4

Observation 0bfa4449-05bd-4707-a3ea-fe2515e5471e · inbound

OmniDrive: An LLM-Choreographed Multi-Agent World Model with Unified Latent Co-Compression for Multi-View Driving Video Generation cites this paper.

OmniDrive: An LLM-Choreographed Multi-Agent World Model with Unified Latent Co-Compression for Multi-View Driving Video Generation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T19:28:52.739321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T01:46:14.430539Z digest=sha256:dad2c8b19da07162eb907827171d8de8e0650da6cd7628945ce14e476c6ac58c

Observation 89427131-e1a6-4272-901f-3792c54795eb · inbound

ReWorld: Learning Better Representations for World Action Models cites this paper.

ReWorld: Learning Better Representations for World Action Models Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-01T18:25:58.424941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T01:58:46.435886Z digest=sha256:a17937c29686b0b356a89968d6b7e0ecbb37e6467cdf2b1ddc6ae027586b89cf

Observation dceccfac-7cfb-4f7f-9062-8aa62f9a58a0 · inbound

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation cites this paper.

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T08:19:04.131379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:19:04.131379Z digest=sha256:f23f275c673d7fcb2e02185dd3a40b81f1e6780210b62e95a6d1b087a2849e7e