Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:23:44.253416Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2606.25360.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:23:44.253416Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-11T15:45:43.483298Z
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fa56e012-6ce6-40a4-b48b-d5e80e006be1 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning π 0: A vision-language-action flow model for general robot control,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 034c7b5b-f9de-4e31-82d0-fc7538c38318 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning OpenVLA: An open-source vision-language-action model,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e1ab2f-2136-4386-9e4c-5fea8e64ece3 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning SAM 3: Segment Anything with Concepts
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 74b78b5e-a78e-4084-aa7a-c55801a491c4 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning ALVINN: An autonomous land vehicle in a neural network,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3239d6f8-2b3f-4200-a385-b3431935b6c8 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bb255f11-f354-47d4-84f7-7d2e69fc3472 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Learning fine-grained bimanual manipulation with low-cost hardware,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a2444a1-2ec6-403e-b4a2-ba672a804eca · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Diffusion policy: Visuomotor policy learning via action diffusion,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d921627-8eff-4cbe-a854-7f02fd358531 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning FiLM: Visual reasoning with a general conditioning layer,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 988e34b6-fd3d-4425-9b4a-a98309529043 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Learning transferable visual models from natural language supervision,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb07ed91-7e93-4a20-a6f1-c0e0ac4f51c6 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2c10252d-d4a1-421f-a607-2e71311693ec · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning 3D diffusion policy: Generalizable visuomotor policy learning via simple 3D representations,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a72c63c-cce9-42f8-88fb-9f599b8d6e4c · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning RISE: 3D perception makes real-world robot imitation simple and effective,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed1ddf3b-764b-42c5-9e64-83c472a43a57 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning RT-2: Vision-language-action models transfer web knowledge to robotic control,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e85e71e3-8618-49c5-a24a-8e2cc70719ec · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Dense object nets: Learn- ing dense visual object descriptors by and for robotic manipulation,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67ef300-7f55-4893-b806-b931d961480b · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning VIMA: General robot manipu- lation with multimodal prompts,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68320909-8301-4fd9-8cb8-0ffbc1d4656d · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning MOKA: Open-vocabulary robotic manipulation through mark-based visual prompting,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ad94462-be3c-472f-a3b8-350a8bda9438 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning ReKep: Spatio- temporal relational keypoint constraints for generalizable robotic ma- nipulation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 596185ce-c585-4a7d-952d-58411ee7f0e5 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning ProtCLIP: Function-Informed Protein Multi-Modal Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1001b5fe-7477-4d1d-80f9-c7768c154381 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Spatial forcing: Implicit spatial representation alignment for vision- language-action model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 26a75245-f23b-4850-902a-b68bd98c7382 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning More than a point: Capturing uncertainty with adaptive affordance heatmaps for spatial grounding in robotic tasks,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5d95502a-351d-48bb-abf4-29b917a635b1 · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning SAM2Act: Integrating visual foundation model with a memory architecture for robotic manipulation,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31ecf627-bde9-4b78-adf1-12745b35a77c · outbound
Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning Deep residual learning for image recognition,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb82373d-f568-47a7-b9b6-aecb5a80be73 · inbound
KAM-WM: Kinematic Affordance Maps from Latent World Models for Robot Manipulation Decoupling Semantics and Geometric Grounding: Spatial Visual Prompts for Language-Conditioned Imitation Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.