Pith. sign in

Paper Citation Record · LEDGER

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation

As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 2 inbound Pith citation observations for arXiv:2606.17598.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.17598 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T00:49:13.291897Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T08:00:25.815355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T10:37:56.081712Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact24
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46797790-d8c0-453b-82dc-85b16fa4f265 · outbound

This paper cites 3d cavla: Leveraging depth and 3d context to generalize vision language action models for unseen tasks.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation 3d cavla: Leveraging depth and 3d context to generalize vision language action models for unseen tasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.822891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:753f83001fef408c7f9d7409de40f3e0bf65a243e6f44f5178cc2f114dadc5e3

Observation e14c2d66-b1b0-45df-a59e-618e141dfc49 · outbound

This paper cites VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.812414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:1ab2e8404900483cd0ec3a09a3f953b8a7054bcba947e6bab7369ffdabdb2648

Observation 6dd5dcf6-22f3-4aec-8b83-9d821755f189 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.756995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:d1bd1612fa8a854604529a75bcb5cfe2e3fe758928fb6e7453ef0ef5b36ebe64

Observation 85314b8f-2b72-4d22-a48f-92cffdf9c9dc · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T21:08:58.819623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:27369a8fc0295b156396253916a8065d5f6b90648cddefa34e81c6e6029657d7

Observation 6f3308f4-a742-41cc-8ea8-0095099d5e4a · outbound

This paper cites SAM 3: Segment Anything with Concepts.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SAM 3: Segment Anything with Concepts

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T21:08:58.817305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:cb0f5cce0b0cfd060ad9bdc0928de40d9a6915d656512fba26a3e7d5dfbbd149

Observation 4b4cb3ff-20d9-4aaa-8d39-55be11425542 · outbound

This paper cites DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.759587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:21ce2141abaaa7188d6965d3e4c7175d8659d037db927c04716ff7cccd60e500

Observation fb6e66dc-dea3-4e49-8fd1-849908118487 · outbound

This paper cites OmniVLA: Physically-grounded multimodal vla with unified multi-sensor perception for robotic manipulation.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation OmniVLA: Physically-grounded multimodal vla with unified multi-sensor perception for robotic manipulation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.750058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:a26864a803122728b78ac8774a56cf7538245be5c2ec0a9c2949be4917d62900

Observation 6efc9740-b259-4526-bb78-3e383b28ed99 · outbound

This paper cites arXiv preprint arXiv:2504.02477 (2025).

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation arXiv preprint arXiv:2504.02477 (2025)

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.828629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:f9c32e3ead8b0ddb29d914af3b70b4033917e0c4aaaea8142bcd0b5735b0a8ce

Observation 04c89fd2-1c34-4d4b-bfb2-66eadae64d45 · outbound

This paper cites Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.819557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:bf06b35640b18904125d9a4f4f76666a865b567e5262b98c2244fd95c257515e

Observation 3a412ad3-eade-46ad-b5fd-968b912f5a90 · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.816613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:1fd81dd3e26a8b278dbb97ea8b7fcf9efc1e03293ce1bc799d25fff6eb8c2e69

Observation 5fe91a17-96bc-4eca-b856-854b34cbc158 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T21:08:58.822787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:8c889ae6b65d3bc279bd60621fbf5fcf4f654b0669050a525fe5bf54e10e93fd

Observation 3e9a22c7-f8ad-4b1f-97e4-184df561b593 · outbound

This paper cites MolmoAct: Action Reasoning Models that can Reason in Space.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation MolmoAct: Action Reasoning Models that can Reason in Space

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T21:08:58.811280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:fcebd1f3768503f30dd3e289599b072c2076f3b44c2add5d648b9b1decadc394

Observation a93787ab-8e2a-4cf4-b98a-433717efa595 · outbound

This paper cites BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.813823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:a8c37c16b206f5e2fd7e45d5ba65280c0b59974a442b09ec1c3984dc375a7ce8

Observation 2ff00401-52cd-4786-8ccd-a4ea929d93ae · outbound

This paper cites Mla: A multisen- sory language-action model for multimodal understanding and forecasting in robotic manipulation.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Mla: A multisen- sory language-action model for multimodal understanding and forecasting in robotic manipulation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.804505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:675ae4e810fbce97ffaf5b4a7093f0cbc57b2beb0b09126defc8184fb5ea6003

Observation 80929dfd-fe70-4c5c-bbc1-8b7f56f7de58 · outbound

This paper cites Octo: An open-source generalist robot policy.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Octo: An open-source generalist robot policy

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T00:49:13.291897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:3f66939c1f8f6c466bb7bdd6db0b26ee51c41d20fb160fced9fd4c5a62ba3c81

Observation 6dd17db8-49e1-424c-aa14-ecd4a8a41852 · outbound

This paper cites Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.807388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:e51ed538ac54d1014e0c2194ddec51ec85818ac8edc96713b6ebaf6411aecca0

Observation ea047546-4cc6-4609-94d9-94c17ea398f8 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.809725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:96ce9103227f60ca2f6201219ec4a61bdf5197e46a5ae5776e3783e758b38e22

Observation 9902a336-631e-40fd-8977-91e42189ae2c · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.814687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:c11b59b1a6c93cdfeaa3130bdee5a1bcf20696936e6c44f9f90f850d4f94fe42

Observation 36827ffb-a7f1-4771-a2ba-85b8dbd480ab · outbound

This paper cites PaliGemma 2: A Family of Versatile VLMs for Transfer.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation PaliGemma 2: A Family of Versatile VLMs for Transfer

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.825380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:f2daf060c8a8931728b5d7e9f6aa554fb3cf6d407ce7c4311e1553e0332b5d66

Observation d0383265-2a3f-47e2-bb6e-205a3d08be8f · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Gemma 2: Improving Open Language Models at a Practical Size

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.787488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:1d824325bfc2fbc21d44e10b4de3046c9a69e6f94362d365856f028a2b5c74c5

Observation 34c70e45-b242-45b0-b03b-e35f68abf9f3 · outbound

This paper cites Tai Wang, Xiaohan Mao, Chenming Zhu, Runsen Xu, Ruiyuan Lyu, Peisen Li, Xiao Chen, Wenwei Zhang, Kai Chen, Tianfan Xue, Xihui Liu, Cewu Lu, Dahua Lin, and Jiangmiao Pang.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Tai Wang, Xiaohan Mao, Chenming Zhu, Runsen Xu, Ruiyuan Lyu, Peisen Li, Xiao Chen, Wenwei Zhang, Kai Chen, Tianfan Xue, Xihui Liu, Cewu Lu, Dahua Lin, and Jiangmiao Pang

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T00:49:13.291897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:f110208aaa5019fbc386342a0cbcb10eafb658bda1535329a1fbeb4b0b8b426a

Observation 3c3b8c0d-17ea-4740-98af-b9410d82b5a7 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.781134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:3110786d8cd5ef24ae3805616c7ba38958913f4918b3763c0e67580b96388369

Observation 1ecde03b-1c19-4ab7-9ece-d491ae0cf1a1 · outbound

This paper cites Unleashing HyDRa: Hybrid Fusion, Depth Consistency and Radar for Unified 3D Perception.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Unleashing HyDRa: Hybrid Fusion, Depth Consistency and Radar for Unified 3D Perception

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.790273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:347217b8eb30f801a48e3caf793b999eb2727b038c024602b9c565d7f902c0a7

Observation e846c0ce-820a-45fb-86bf-aa3999ff70bd · outbound

This paper cites A Pragmatic VLA Foundation Model.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation A Pragmatic VLA Foundation Model

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.803528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:bd34cf92ce6f0190c30a8e21907c28c5c1add4906e05066b3bb4aaa17cce8ce3

Observation 5d816bc4-fa2f-4d66-b12a-1c94e094fce6 · outbound

This paper cites Forcevla: Enhancing vla models with a force-aware moe for contact-rich manipulation.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Forcevla: Enhancing vla models with a force-aware moe for contact-rich manipulation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.792833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:312445e5a66d8c6d0c7b9625d828ec5f2aa82887dc03e7e5ff54bd12ad8a87c6

Observation ed59098a-cfc3-4986-b4a1-1da622e0bebc · outbound

This paper cites Generalizable Humanoid Manipulation with 3D Diffusion Policies.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Generalizable Humanoid Manipulation with 3D Diffusion Policies

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.774480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:afdfe9c4d086fb4ea24bb2589514b1ba80fecdd265c7c2039ebb7ad8367ab232

Observation f0cd5287-2188-4a2f-8774-2c27d031698b · outbound

This paper cites Clap: Contrastive latent action pretraining for learning vision-language-action models from human videos.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Clap: Contrastive latent action pretraining for learning vision-language-action models from human videos

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.776982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:94294529f1c104f8d8b330a3dd64a45c2b64e2c21f805b7aa66be00a1e40047b

Observation 12c552b4-9e30-47a5-9f6e-02907289948b · outbound

This paper cites VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.769029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:19fa16e6708dc829a391d0d7caa064d216f6b4b7972d711266b583245442f6bc

Observation d053fa15-d9e8-4290-b91d-136925662772 · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-03T21:08:58.765321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:28bd2d37f1c5144f2cafa43fc0ca611aa7f24d0d2d76c644b82a2eb254bd7fd3

Observation 959b0368-ad92-44e5-896c-1ce841bf1b05 · outbound

This paper cites Doracamom: Joint 3d detection and occupancy prediction with multi-view 4d radars and cameras for omnidirectional perception.arXiv preprint arXiv:2501.15394,.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation Doracamom: Joint 3d detection and occupancy prediction with multi-view 4d radars and cameras for omnidirectional perception.arXiv preprint arXiv:2501.15394,

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.769127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:acb218f3428fd89e76a697adf85946d65255a0f4aa2d4cc3c61122167c6ed9a9

Pith citing papers

Observation c3869826-5bd1-4b4e-88e2-fc38f7b9a64c · inbound

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots cites this paper.

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T10:37:56.083179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T10:37:31.925624Z digest=sha256:457668306b40d4db35706d4541f860797fd351200982f8bd52a956147f76fc11

Observation 061015d8-8224-4e91-b492-b5e23be24f5b · inbound

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots cites this paper.

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T08:00:25.815355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:00:25.815355Z digest=sha256:85e6f9c3fd82d6b32ecefb580d8dfe043c759077d8d62f1800f4f3f19c31db32