Pith. sign in

Paper Citation Record · LEDGER

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

As of 13 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 8 inbound Pith citation observations for arXiv:2508.07650.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07650 v2

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:59:04.663968Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T17:38:20.785896Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 0a654059-9e67-4146-9782-73063cb25d2a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.281025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.281025Z digest=sha256:cf9c919972d745d5b4674c7f356ad1d16827eaeabfd14359beba5d9b3ead67aa

Observation 633a45ab-0d9e-4ef9-b84a-796c397622bc · outbound

This paper cites write newline.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.374741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.374741Z digest=sha256:c099e9a5869a317d9bf4ea4aced307b8a207e50863204d8f037fe4034ade3bc5

Observation d5704008-07ae-4823-999f-98e7addf7ddd · outbound

This paper cites Qwen2.5-VL Technical Report.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.513385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.513385Z digest=sha256:d39e3f2cabed0db89d35d55d8e7bb97204ef79a730b990a6f292a9b9aee7225f

Observation 4b160908-3609-46e3-bba5-8edf69cc572f · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.642456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.642456Z digest=sha256:ef0eb6be7c14e76c2d57671be938179338483af6d2f7f49f58a02188de9717d3

Observation 1559dab7-3594-454f-bd9f-302cac670c54 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.765491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.765491Z digest=sha256:5d5b25b118112978266d0a6fdafd7f321fb6384a6fa63e14c919569644dec110

Observation 2ab685a4-4b7f-4a09-b1be-3b9db0277c27 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RT-1: Robotics Transformer for Real-World Control at Scale

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:00.915445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:00.915445Z digest=sha256:00358cec0309e1d3365888ef0b5c573fe5189581ab2bf5bc0a3eafb8695991eb

Observation c99be82a-8dec-4b14-9ead-6d31a1e9b659 · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.047936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.047936Z digest=sha256:627a89d95d7e5b0c301db96878a06ec23c85565266b0ffc2ad98a4a7bbe1cc13

Observation 20ff8100-4ca7-47bb-9441-8c965fb14dc3 · outbound

This paper cites Training Strategies for Efficient Embodied Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Training Strategies for Efficient Embodied Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.209286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.209286Z digest=sha256:699d6c67c15c4526a8925161e3f9760cda654e8c7faab7301c3145e6dfaacb4d

Observation 6dd3773a-3739-4d0c-ac3d-15d3faa5cb56 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.316143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.316143Z digest=sha256:d7a1709ed9bba8e01b7a94cb4677803a0bc84644fc8010117b9304edaf1419c2

Observation c0535c25-b455-4093-be57-bbe23e6ed593 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.465144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.465144Z digest=sha256:db13ed1ff5c51d3b84244cd41729f123eabe07d9a1019696be774f6a62608645

Observation 8112ce35-afb6-4377-be70-593bac236238 · outbound

This paper cites U.; Akram, W.; Saoud, L.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions U.; Akram, W.; Saoud, L

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.586486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.586486Z digest=sha256:cead4c17c87c65821e275f831a7cf38e34585b693533208982e3cb01ed39a502

Observation 9c7419bf-4469-49a7-aeb0-36cf0ae16ef8 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.753969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.753969Z digest=sha256:6bf1430b03033cc6e6ef65839b3a4b0cdd649fb00c8047bedff91605ce35460d

Observation cbf66c8f-7dff-4bcf-af3d-5e294652bcc7 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.879712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.879712Z digest=sha256:d8a3585fc04ec5a3cb1a645e3a4dfc5adb214a34d194b8d0e19dab64339e0fb2

Observation f22f2f3f-3975-4f85-88c5-a93078de177f · outbound

This paper cites Flow Matching for Generative Modeling.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Flow Matching for Generative Modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.984771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.984771Z digest=sha256:14fbfec0fe10f6cbfa6f37a8f63566579f86d22739280887b61e373123217759

Observation 21b237a7-900b-4669-83fe-c4bb2e4577ba · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.125016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.125016Z digest=sha256:306c19bfc44aa366620e48299ca81303b8339701c7558a1bdc0c9f12c17ab680

Observation d2a5271a-649c-47e2-89a7-f093f194ff20 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:06.090482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-05T21:59:02.219607Z digest=sha256:cfc041ea8e546588cb337269b884f7b938d27e2403e24e9c573349062cea3c7a

Observation 2d951652-97cf-4845-b9e3-30536abadda5 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.395778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.395778Z digest=sha256:314a38bd8cd6004f17b8dd7181fadfb72aa079fd9cdb84a8af7dfd827fad11e4

Observation 3972e8e3-e061-4e2c-a015-369cad400f00 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.521887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.521887Z digest=sha256:de26df0d14b3d5a494ca5466be3b7286350a4c777f154066828c478275d1319e

Observation 1d9a072c-0eda-4c4d-b4ff-bd99563f9f45 · outbound

This paper cites C.; Hagenbuchner, M.; and Monfardini, G.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions C.; Hagenbuchner, M.; and Monfardini, G

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.669629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.669629Z digest=sha256:e939f2d17ad3049f82a73454c637c19b36b63e28f4e1a988c70d5088ee79567a

Observation 387a2d33-f234-4fb8-8dc5-7b234f572307 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Octo: An Open-Source Generalist Robot Policy

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:02.850855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:02.850855Z digest=sha256:f602df878c2cca5af8cf95cc1601af0049fec28ced35eec2b034589e20903d85

Observation 1eb08441-3332-416c-85bb-0788f87d9060 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions N.; Kaiser, .; and Polosukhin, I

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.028529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.028529Z digest=sha256:812f790fddb09d2eac445efef1df2defb646d21e866dcfd5bdd8c1f4be5b8458

Observation 46b3d570-56a7-4917-9861-7454eaa9e4de · outbound

This paper cites RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.238165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.238165Z digest=sha256:1c7976ecb736246e6bb5808ea0ff394744edccc40dbfaf814a985fbcb35e6a99

Observation ed56506c-f6b4-4bdb-9e27-8837f393f21b · outbound

This paper cites V.; Zhou, D.; et al.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions V.; Zhou, D.; et al

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.360868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.360868Z digest=sha256:91d9043824c7a03ed459a0b7de90bb96f2db88de47e90d7f2c1a3418c4ecb0a7

Observation 8a3068a2-2413-4b17-893b-59a5f87d8200 · outbound

This paper cites Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.506693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.506693Z digest=sha256:d1e29181a9d7a5bb179ed40fde0e48c3a3fa71f8b7fe5b101d91ac679aa26917

Observation c37b3efa-6603-4150-a9d7-3b8038644537 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.685804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.685804Z digest=sha256:8a94ba531760cb49fbb7b73eca32cd200d0e2c06aa81b543582739cb8cc0d2fe

Observation 937d58e2-100c-4b34-8746-37cb3ae1f068 · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:05.824403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-05T21:59:03.815934Z digest=sha256:bdaf06e34f19c135cc9fb264fd84a42365966708aeb80a40d3338f662f265618

Observation 607ad875-3aa4-405e-8f2a-2638e625de80 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:03.989579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:03.989579Z digest=sha256:1a3edeaa55e667734e192be3cff1d6da84cc2847f6754cc2b3a776717dcd1a76

Observation 10fcbea8-bb85-41dc-8eb6-083a24f9dbee · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.110925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.110925Z digest=sha256:65ffcc40f0fc6707783763adc6cee3a37d019000f6eeee23f7470fc3d72ac266

Observation fc824915-7486-47f8-a96f-94bace06dedf · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:59:05.549634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-05T21:59:04.244235Z digest=sha256:b024f5583bafa96fdfebeaacac51687c3b107b8fcb56c271bbee00e2d4bc71ba

Observation b6b0a23a-ae10-4209-9a43-49c754fc910d · outbound

This paper cites J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; et al.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions J.; Fu, Z.; Zhang, Z.; Wu, Y.; Li, Z.; Ma, Q.; Han, S.; Finn, C.; et al

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:59:05.235334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-05T21:59:04.376600Z digest=sha256:81301dc490cbc22a4fcea355eec9aa0d3a75250bba0e3ac80bf82764b99082a7

Observation a04cce51-79b9-43fb-ada7-aafa8657335a · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.551521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.551521Z digest=sha256:728bf01fb2b7e818dcdbfdc9a59b4795df1ec0fd20658b5b4bb9671881ca34d7

Observation e7a23773-4d3a-433b-bee1-e3af492139df · outbound

This paper cites an unresolved cited work.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:04.663968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:04.663968Z digest=sha256:fe3fe96b58fd17540e5012d97537f5d89fcb1298c12e89ce45e84c176a861bf2

Pith citing papers

Observation cfaa48a2-6d08-4c1e-8e5f-e7a9f4ae9992 · inbound

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer cites this paper.

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T17:38:20.785896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:38:20.785896Z digest=sha256:026e328f3905de78bba632519851258f44b84ff04561a8ef89169ac6ee8c104c

Observation 6f36ca5f-1c77-4b7a-acf9-74fb39086d8e · inbound

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models cites this paper.

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:57.258483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T16:36:19.995405Z digest=sha256:a402b62f1d3d8849ea46775e3e7742b6c1a3c598d5fb69f32f250265aae4d452

Observation 88d240b7-7eaf-43ec-a10e-27dd97ffe09f · inbound

TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation cites this paper.

TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:09.179305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-08T14:54:48.058400Z digest=sha256:76ecf9247c214db932157b73ecc316a977f6b65ba37889206aa05fd4208d3475

Observation 7b5158b3-e675-40d2-8a40-aa4907ff9396 · inbound

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction cites this paper.

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T04:52:17.025802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T04:50:07.792757Z digest=sha256:d860d0b822d0277df7049799298f94058190b2111952a23f64b797946dd6e424

Observation 5f93b8ea-0f8e-4842-9f73-437ff3206344 · inbound

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data cites this paper.

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:57:33.311170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-14T17:54:50.325820Z digest=sha256:cab2b197a5301f56bac488e98b0d4001b2a757ce1c05a0053341a7df0ae4ce86

Observation 963a8b05-0fcc-418c-907c-038d6af2d935 · inbound

DSSP: Diffusion State Space Policy with Full-History Encoding cites this paper.

DSSP: Diffusion State Space Policy with Full-History Encoding GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:21:24.247338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-22T10:17:44.804745Z digest=sha256:2780d62a58c182a1b5c4c27373bc49a9f2e0d691a23ab680b4aa909a2b56f05f

Observation 8af57b5b-6e93-4cc6-9456-acc96e48fe0f · inbound

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models cites this paper.

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:01:20.749485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-27T21:57:19.195413Z digest=sha256:03086ca7c263ce7e5a30f6dc03e94744aecda602fc64d5b64d56d1ae01b58fbb

Observation 4a2bf18c-6494-44e4-b6a3-ad0de77e6d30 · inbound

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation cites this paper.

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T06:36:27.368310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:36:27.368310Z digest=sha256:70622eab782652c5e3e3f2d6355ade90760b927d2c82c851b6d881013edee950