Pith. sign in

Paper Citation Record · LEDGER

DriveVA: Video Action Models are Zero-Shot Drivers

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2604.04198.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.04198 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T18:25:51.336494Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T15:49:56.319991Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 59418f40-5ca5-4617-8516-11e7b923cfee · inbound

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model cites this paper.

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model DriveVA: Video Action Models are Zero-Shot Drivers

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-22T07:41:15.002227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T07:37:11.292270Z digest=sha256:38ff953eeaef9eb9f98a9b0b9d6970d8a0168bb1de98526c5bc550d30136fd33

Observation bbb9dacf-b393-4680-a758-6e83f2130ffa · inbound

World Action Models: A Survey cites this paper.

World Action Models: A Survey DriveVA: Video Action Models are Zero-Shot Drivers

Reference 106

Resolution
verified exact
local_arxiv, observed 2026-07-04T04:09:35.437441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T17:11:12.686936Z digest=sha256:d1119ab8362742a0c6807722fe4db414465c77eb56128c02bd8bd972b43957c6

Observation cf9a880f-808f-4e6c-90fb-ca1bbb819704 · inbound

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models cites this paper.

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models DriveVA: Video Action Models are Zero-Shot Drivers

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-07-04T15:49:56.326328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T01:27:24.402364Z digest=sha256:acb411066b52df21bfa6a41f11e65a4dafbb8fd688ab80d53a60ac7577207571

Observation 62bcd138-de57-4696-9d91-c67d7d7afedd · inbound

ReWorld: Learning Better Representations for World Action Models cites this paper.

ReWorld: Learning Better Representations for World Action Models DriveVA: Video Action Models are Zero-Shot Drivers

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-01T18:25:58.419741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T01:58:46.435886Z digest=sha256:6f93f9b6350dcd51d9b0bafba2d79a76788dcf3d1de98937946013d9ab791a6c

Observation 9558e003-127b-4863-bed2-d17b1b7fa256 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T18:25:57.822702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T02:03:45.564122Z digest=sha256:26077226956d2f4ecb32dad659927f7c98b0da5e49546c7b02a8f162a44844eb

Observation ec780c94-118b-402a-81c8-e5362ef57eb2 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:35:40.605408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T06:25:58.872140Z digest=sha256:92ef5799fe4fbf156cdbfd66998dc4257feec15052fe396bee23b3d1fdc4ee01

Observation 019dce4a-5ca4-4039-abe0-e582d0ca74e9 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:22.804185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-02T20:52:28.444524Z digest=sha256:67a2e870363709fdd3b8f03040f54465aac8ae82fd6e2585717828b49b60fa60

Observation 52d8ebee-90c3-4074-ad80-4db284759124 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T22:49:00.854031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-03T22:44:16.272541Z digest=sha256:0cfc0fb8ae002f4e9854257142ed57b7ad6cc831baeabdee0306ad40cadf6f12

Observation 49217cb7-29f7-4119-b24b-0dc8c8adf95c · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-14T17:14:19.770867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:14:19.770867Z digest=sha256:9a3eb831b0c249a2f1e0a28dad087b661caeeacd39cdee7d51d6509f6af2aa94

Observation e6ec0803-aa07-4b6b-a7e7-4654692134ab · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments DriveVA: Video Action Models are Zero-Shot Drivers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:12.168811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:12.168811Z digest=sha256:389092279ae516c9d818487beed9869c4d26cf830420514c42eb32f6312bd736

Observation 3453d33a-5041-492d-a05b-09dd659a81d1 · inbound

Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning cites this paper.

Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning DriveVA: Video Action Models are Zero-Shot Drivers

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T16:38:39.653582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-03T16:37:06.384435Z digest=sha256:2d04e929df8fd2418d2bbdc500b8f41b14388d6b046be44cd17945d7b2584170

Observation 07079a8a-bacb-4af4-9869-872674882e70 · inbound

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation cites this paper.

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation DriveVA: Video Action Models are Zero-Shot Drivers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T08:19:04.131379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:19:04.131379Z digest=sha256:0ae3d16e7dd853b4c9dd7964f3d906faa8281d76ecb9f07a692ae6f59d825f2c

Observation ac254dee-5964-4374-867b-ae10eb1698ec · inbound

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features cites this paper.

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features DriveVA: Video Action Models are Zero-Shot Drivers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T18:25:51.336494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:25:51.336494Z digest=sha256:b09a1cd948a7d0f4f86378fbde2d8b26bae08220e88ea066309802c77475dc92