Pith. sign in

Paper Citation Record · LEDGER

VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2502.13508.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.13508 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:07:25.936596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:30:00.525942Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 387c0b76-0387-4d93-8362-cc628d7755df · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.994403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:9e90ed8240150a99543f40cfd30f894cb020398d41d1a84f3d6851eeb283c1e0

Observation ef8e333e-3baf-4787-8619-ad432f15f447 · inbound

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers cites this paper.

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:25.936596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:25.936596Z digest=sha256:573fecdb5e47365152cc342a626544a92bb68a2488104c48f54258b67c40bd67

Observation 5e1a6cb3-5c40-453d-b33e-088a8634317a · inbound

Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive Prompting cites this paper.

Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive Prompting VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T14:34:25.882971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:34:25.882971Z digest=sha256:3a8bfdd9ab8176d8916e7a998c4510deecfa72c5c7fc9a81988ca636fdd655ac

Observation e7d6c07d-6368-4be3-8a11-877a77ba460e · inbound

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation cites this paper.

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 149

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:08:02.092746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T15:05:21.907878Z digest=sha256:46f82253b3851485e0f64b90492cb65c237c6f13d28f97d9e612c51f8afc8f6f

Observation 5bdac23c-6650-4273-aff4-9e98bbd7a717 · inbound

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making cites this paper.

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.841267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T00:46:03.360770Z digest=sha256:57a19eeee6184ed945748457bd7e9c74ba14361ba26e19f2a6e61fe203ee01c0

Observation 31c068bb-f719-4f21-997a-637e5a176494 · inbound

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model cites this paper.

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:09.582834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T11:45:18.081248Z digest=sha256:343316513383305f5910a238906055cf9c55d1065f377a65926792e6880f9d5b

Observation 34a23613-c686-48b8-b1c4-34a9376162dd · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:25.676152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:08:43.222818Z digest=sha256:83b68b8711876e17206407529f6a7d154e82207001cb92890be262970e4772db

Observation db8ae8a9-2adb-461a-a5bd-d466c803c1e2 · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T14:25:59.725499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:25:59.725499Z digest=sha256:c7075519c4f9db6f128198949361fddd45683b6cb1feee1d6ac09795673c823b

Observation 12c552b4-9e30-47a5-9f6e-02907289948b · inbound

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation cites this paper.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.769029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:eec9fcbe331e5b08e74ee1d63b2dd958198bbd6915dc86ba010e55b5ff60b159

Observation d9c2a461-1dfc-4374-b928-9b2f0ea7cf50 · inbound

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation cites this paper.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:30:00.527553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T23:39:14.779122Z digest=sha256:ffdb2488f6ac392bf71c620a8b85eec6933af9383c6b872dce7486d5ea8024ea

Observation d0f4b36f-b5a1-4bcb-ae7a-03fd37fc19de · inbound

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation cites this paper.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:d8299a954de907f19ed3a7376558219cb73bfd56465cb84e0db0977d34412812

Observation 2cda399c-9843-416e-91da-f3a99b59713c · inbound

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation cites this paper.

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T06:39:16.303751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:39:16.303751Z digest=sha256:002253b6f3f799259a5d69b43bd0efeeafc5a94b8c7d78ddacd39604b5168d94