Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T00:49:13.291897Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 2 inbound Pith citation observations for arXiv:2606.17598.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T00:49:13.291897Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T08:00:25.815355Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T10:37:56.081712Z
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 46797790-d8c0-453b-82dc-85b16fa4f265 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation 3d cavla: Leveraging depth and 3d context to generalize vision language action models for unseen tasks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e14c2d66-b1b0-45df-a59e-618e141dfc49 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dd5dcf6-22f3-4aec-8b83-9d821755f189 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85314b8f-2b72-4d22-a48f-92cffdf9c9dc · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f3308f4-a742-41cc-8ea8-0095099d5e4a · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SAM 3: Segment Anything with Concepts
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b4cb3ff-20d9-4aaa-8d39-55be11425542 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fb6e66dc-dea3-4e49-8fd1-849908118487 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation OmniVLA: Physically-grounded multimodal vla with unified multi-sensor perception for robotic manipulation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6efc9740-b259-4526-bb78-3e383b28ed99 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation arXiv preprint arXiv:2504.02477 (2025)
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04c89fd2-1c34-4d4b-bfb2-66eadae64d45 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a412ad3-eade-46ad-b5fd-968b912f5a90 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fe91a17-96bc-4eca-b856-854b34cbc158 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e9a22c7-f8ad-4b1f-97e4-184df561b593 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation MolmoAct: Action Reasoning Models that can Reason in Space
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a93787ab-8e2a-4cf4-b98a-433717efa595 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ff00401-52cd-4786-8ccd-a4ea929d93ae · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Mla: A multisen- sory language-action model for multimodal understanding and forecasting in robotic manipulation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80929dfd-fe70-4c5c-bbc1-8b7f56f7de58 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Octo: An open-source generalist robot policy
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd17db8-49e1-424c-aa14-ecd4a8a41852 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea047546-4cc6-4609-94d9-94c17ea398f8 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9902a336-631e-40fd-8977-91e42189ae2c · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36827ffb-a7f1-4771-a2ba-85b8dbd480ab · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation PaliGemma 2: A Family of Versatile VLMs for Transfer
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0383265-2a3f-47e2-bb6e-205a3d08be8f · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Gemma 2: Improving Open Language Models at a Practical Size
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34c70e45-b242-45b0-b03b-e35f68abf9f3 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Tai Wang, Xiaohan Mao, Chenming Zhu, Runsen Xu, Ruiyuan Lyu, Peisen Li, Xiao Chen, Wenwei Zhang, Kai Chen, Tianfan Xue, Xihui Liu, Cewu Lu, Dahua Lin, and Jiangmiao Pang
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c3b8c0d-17ea-4740-98af-b9410d82b5a7 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ecde03b-1c19-4ab7-9ece-d491ae0cf1a1 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Unleashing HyDRa: Hybrid Fusion, Depth Consistency and Radar for Unified 3D Perception
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e846c0ce-820a-45fb-86bf-aa3999ff70bd · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation A Pragmatic VLA Foundation Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d816bc4-fa2f-4d66-b12a-1c94e094fce6 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Forcevla: Enhancing vla models with a force-aware moe for contact-rich manipulation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed59098a-cfc3-4986-b4a1-1da622e0bebc · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Generalizable Humanoid Manipulation with 3D Diffusion Policies
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0cd5287-2188-4a2f-8774-2c27d031698b · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Clap: Contrastive latent action pretraining for learning vision-language-action models from human videos
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12c552b4-9e30-47a5-9f6e-02907289948b · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d053fa15-d9e8-4290-b91d-136925662772 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 959b0368-ad92-44e5-896c-1ce841bf1b05 · outbound
MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Doracamom: Joint 3d detection and occupancy prediction with multi-view 4d radars and cameras for omnidirectional perception.arXiv preprint arXiv:2501.15394,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3869826-5bd1-4b4e-88e2-fc38f7b9a64c · inbound
Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 061015d8-8224-4e91-b492-b5e23be24f5b · inbound
Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.