Pith. sign in

Paper Citation Record · LEDGER

A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.17418.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.17418 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:23.963963Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:05:45.043767Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e0f07e5a-4754-434c-b3ad-c1f867011397 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:00:48.968379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:fe660f6c24cdf4cdee9f2a0d9b821195d5619e83103ab13523ea309626eb1a3c

Observation 2cc89cae-e33d-429f-b495-13b8f6e381fc · inbound

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models cites this paper.

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:23.963963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:23.963963Z digest=sha256:76522fbe873080dac4f1c543b392aee26bd4b48f50b42fe6b159eb6f7f7b0fc0

Observation 2f0ebc89-1577-484e-b049-91abb1db6d05 · inbound

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning cites this paper.

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:34:16.484483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:34:16.484483Z digest=sha256:70145135a6a1507bcbff596b9c9c5ad5957a48a9e2478e1d385568a32d359db0

Observation 5c472cac-20d5-4e4e-a927-03706a803155 · inbound

Reinforcement Learning for Flow-Matching Policies cites this paper.

Reinforcement Learning for Flow-Matching Policies A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:41.911320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:47:41.911320Z digest=sha256:549055e8d79f5c90d2de56c5155c12eabdf589c9326858add59b1072cb1b9226

Observation fb69eac5-e136-44f9-8330-3e71a657ff1a · inbound

AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models cites this paper.

AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T21:30:18.227523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T21:28:18.630934Z digest=sha256:c4b54a3b802d016b3bbdf3906f6d52f5fd1c714840ce9c72e74de0ed498fa3b9

Observation f6d0d885-b373-4912-8df6-9478e458394b · inbound

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding cites this paper.

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T12:37:27.492757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:37:27.492757Z digest=sha256:0df3c7d9e4a3474c06a87abbea4709d65c3c1ab4a039714604fc88217ab06b5d

Observation 3243e82d-d122-4007-a230-2589704edc41 · inbound

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos cites this paper.

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T22:24:08.310224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:24:08.310224Z digest=sha256:b139f835deb99db3e7d109ca85d5c2e73c65621f335ac692e033061186f97d6f

Observation 241cb840-a15f-4c70-aad4-53589ef4e32a · inbound

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery cites this paper.

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:46:04.903328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T15:14:03.865035Z digest=sha256:6a1ba09706f65cb9d19602e9a6ac22b34809a867aef9fb4583689da3c35b1dd3

Observation f9291fae-bfa5-4287-87a8-e78dad21e8f8 · inbound

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery cites this paper.

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:05:45.046082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T01:10:48.111387Z digest=sha256:f9d6bca8833c087e4bf9f0c50b659bbb827440c930427d7cef11133deb1a52fc

Observation b3f8ac2d-30d1-402a-824a-f9a9eea32c92 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.367232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:bd9bba01a6fe861e4a046e05367bf8047c10d5f4abe8e347cea02da0345b7e2e

Observation d92c9ae3-9b50-4c62-bab0-2f759da5d721 · inbound

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots cites this paper.

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:21:31.544358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T05:17:03.803350Z digest=sha256:5d82423ac9e2e094558a6bada3cb4ddf028d8f80a4856c5196f5aa06b56ae31d

Observation 54ab0ffb-ebc6-4b14-8e26-05d8b4256dd5 · inbound

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots cites this paper.

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.871897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T22:24:25.871817Z digest=sha256:1c1f884722af1f979b4b7924dcfa5ac6a463e19f147997780ba5d278d973c09d

Observation 1cc93585-4968-42d8-88b2-8af218133987 · inbound

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling cites this paper.

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:27.701922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T13:33:03.368006Z digest=sha256:1f7926d18f963a5934815fda1f85708e9da203f485f46ae882f183fc20fb62e1

Observation 839ce8f3-afe6-41b9-9a6c-2f203e529d79 · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:56.910815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:56.910815Z digest=sha256:ba55bf37f4e8e0e6c35ce33df36cb66c2faa19891ecafd8a8021559589ed1340