Pith. sign in

Paper Citation Record · LEDGER

Generative Image as Action Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2407.07875.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07875 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:51:27.058953Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:29:51.759974Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2eb005ad-2e6c-4066-9b97-2d26fc012eb2 · inbound

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models cites this paper.

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models Generative Image as Action Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:21:45.023674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:21:44.903048Z digest=sha256:3f5adadcfb80cd9b3978c91acd7268c05dcc5afa373419d77049de25b5f5cabe

Observation 7470c18c-0677-4a94-9b1c-7e7350a54004 · inbound

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt cites this paper.

Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt Generative Image as Action Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:27.058953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:51:27.058953Z digest=sha256:69ea872e8991c88e9c657a77b655f6dac73b7b34faab36bd113f5fbc1451d685

Observation 32643211-05b3-4b02-a9cc-5e80cf32bac8 · inbound

GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation cites this paper.

GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation Generative Image as Action Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:50:29.339711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T07:48:33.832017Z digest=sha256:9f2c0823ad4f566e1f4a001ddde01386211157f5c0bceca3e7442c67db87b310

Observation 4a48fe91-d48d-461d-91c0-23f106df79d9 · inbound

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges cites this paper.

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges Generative Image as Action Models

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-05T16:55:52.447713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:55:52.447713Z digest=sha256:a0fe2aa08a652db3b2948545b659784f90baeb5ce90646dd4c43b8ebb141e7ef

Observation 8072a741-9bb4-46c7-8fdf-617405354e20 · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Generative Image as Action Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:42.288905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:42.288905Z digest=sha256:ae07000bfbde7eae7b3a8edbae01a502e9c4adabb22940e2ffe4f7146db09312

Observation 0b2b78df-9ff5-407a-af30-aa908e499087 · inbound

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data cites this paper.

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data Generative Image as Action Models

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:03:00.993184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T17:02:18.358675Z digest=sha256:b7bc0b58eecf5c3a8abdf89e9e6a19e30ae01c8ba207f76996a3d17c06fe267a

Observation bba91838-5247-4504-904f-b4602f1e8b9e · inbound

EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields cites this paper.

EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields Generative Image as Action Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:51:06.485236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T13:50:37.638842Z digest=sha256:b3ff45858cd9b2c9d15c84ff4a6a9ac038ac8f8281cc30a80929a0cc59b85172

Observation 263e24a8-255d-4081-9589-aad58819a567 · inbound

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models cites this paper.

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models Generative Image as Action Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.761408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:10:09.186699Z digest=sha256:959deb45b56241731bac0ec5b4fa070fe0a526de34fea788c69099be7a03bc1a

Observation 1b5e6807-43c8-4859-9179-4dba073b92f2 · inbound

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models cites this paper.

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models Generative Image as Action Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:54:34.859026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T09:51:54.464950Z digest=sha256:ebdfb09e7c288a6394fc487dfeac95874965d9b488140dd4ea08e614bc23390a