Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:19:50.356389Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2608.04633.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:19:50.356389Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 26f9b6b8-c9c3-4892-ab74-7105435e5b9f · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Zitkovich, T
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f472e05-b48c-4833-8843-c2850544695d · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b922b83f-7d15-4061-8911-9e7cd6d05b4e · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e68b6461-d7a1-4051-8dcd-6f2b3aea963b · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5194fd60-de51-48a8-aec8-e97196c28576 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db459a88-3610-46fd-9aaa-6187dbaf2234 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cd5a458-dec3-4ee2-b750-754efc5b8c04 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Singh, A
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92770e72-495c-422f-923d-8c5a48bad184 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e9efc7-a1bd-46c6-8d49-60b75febcc95 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab78dc4e-908e-443f-8b81-016d8a0891fd · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e5c079e-cc74-487e-944a-c312cd13ca68 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f34d831-2310-4095-a423-78be06bfecb4 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 677ce944-fbae-4bbe-babe-38852f0f9f5e · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6238bce0-fabf-42d2-a31d-6f4a6730eef1 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def2df4a-59f4-4fa3-9f90-ec4a4a19695f · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models O’Neill, A
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 567b9100-3b8f-489b-9b99-5978f5f252dd · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a909a4b9-f254-43ae-83a7-e22beb324557 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29970e42-475e-44e3-9b81-74697dd96599 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e3fb0a-6958-4a39-a54f-fa6f9591b62a · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d55763d2-e79f-4c5a-aa60-919c89b1877a · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Shridhar, L
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5947d0b1-08dc-443b-a0ca-54017e52ee19 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Act3D: 3D Feature Field Transformers for Multi-Task Robotic Manipulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43fa7b85-4d46-44ac-9aa4-378678dda7dc · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Goyal, J
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3da0a491-2e47-4f1f-95f8-48f090164ce9 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models RVT-2: Learning Precise Manipulation from Few Demonstrations
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1e07334-0fca-48ef-b91e-2331ec311f7d · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1fd0a23-d6f6-41c7-a9e3-d7f7e487a763 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Any3D-VLA: Enhancing VLA Robustness via Diverse Point Clouds
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 752060ec-438d-4d3c-ad9e-06c242fde276 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c046748c-8091-4aee-bbdd-93269d49b7a9 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc182fc0-0a71-483b-a2dc-10046dc23220 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25f249e9-f40e-49e6-b0ef-f99092c85327 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 767aec74-c80d-4bca-8b52-692215445dc4 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8669714-812a-40b7-8883-a0c45c288110 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd9d14f8-9612-487f-8ec8-2cde133f0f58 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Peebles and S
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06563ac3-47a8-4429-baa6-b30e995206c1 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Radford, J
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d672b3e2-c805-4b58-81a8-5925afb5e535 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Denoising Diffusion Implicit Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc9f1646-a4c2-418a-b34c-61d6444f36c1 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23102af3-4a0e-486f-b583-ba540c05c04d · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Rombach, A
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c149296-e834-440a-89e7-1cfb36079b83 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a4a7068-291a-4154-a9ff-248e47bae85a · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37bcff39-8ec9-46f8-838f-803038cf6f9b · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d46dfc98-2a71-4997-be37-2ba63028c1af · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef3bd510-365b-4b64-a3ab-bce873a810f5 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c004a486-b24f-41c0-8c5f-aa26d8d09415 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58340e5a-1563-48b3-8529-a965af4e892c · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 479ab36c-5a63-4a51-90f5-cdbe56196a57 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e930a4c-d7e0-4273-b5c3-18d9d8fce342 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models Kirillov, E
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c57b9e74-d5e7-4767-8b1f-30d0aa36d0b2 · outbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.