Pith. sign in

Paper Citation Record · LEDGER

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

As of 4 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2605.31286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.31286 v2

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:58:48.611343Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T20:46:04.923901Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact31
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d31bef59-942d-4b01-b1d9-40da007ba36e · outbound

This paper cites Qwen3-VL Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Qwen3-VL Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.394231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:4b49b4b27e91bcb0892a05b361702ca13c2aed4012666cec2af3cd5514b7e346

Observation dc6a7625-44a6-42e1-9a5b-c858f355c4d5 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation PaliGemma: A versatile 3B VLM for transfer

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.428612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:17f39195e5db06b398b611e1f9651e604c4b3b9eb7936254bfb1144a05fbf076

Observation d86c246f-5270-4827-81ed-4d82d80fa67e · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.340717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2fbdf19335a6dc32f144a01387b41b37d75bab937d38d177251ff04b66dd9c61

Observation 86a0d7f4-aa26-40ff-a8cb-1ff68843b1dd · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T19:56:10.423826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:86381416d9633303bbba5d11f6855b67ba93c5ba1d4b10ed9f27559a00228e14

Observation b8b71efe-b976-4e16-8f61-badf360c53e5 · outbound

This paper cites Training-time action conditioning for efficient real-time chunking.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Training-time action conditioning for efficient real-time chunking

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.412825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:75cb8916ba287dc1633e24ddd5faf2a999b7cdcf5aeda4949113cc240d763b66

Observation 3baeb24a-5909-4a9a-9bb7-c2ffca01c367 · outbound

This paper cites Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Real-time execution of action chunking flow policies.Advances in Neural Information Processing Systems, 38:33383–33407, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:c2ef7b9f64e138fdb4f8eeb43c0aed55b84fbd6441f41e5a564a0864d3217265

Observation 0bc8d3e1-d01f-4537-93fe-2d7e6d1492b5 · outbound

This paper cites arXiv preprint arXiv:2602.12684 (2026).

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation arXiv preprint arXiv:2602.12684 (2026)

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.421785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:11fbb544cec59167a912fecea3b65998d78a6061396b8fb9f7afcce10934c407

Observation 3b44f106-5c57-48d6-bd86-8313d80eb5d9 · outbound

This paper cites Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Interactive imitation learning in robotics: A survey.Foundations and Trends®in Robotics, 10(1-2):1–197, 2022

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:f82fb7deb31eec565983a396321eced62200d776d8d73b6b74c77033aa200467

Observation 649ee648-8e63-44a7-8195-f24c1d4bb078 · outbound

This paper cites RynnVLA-002: A Unified Vision-Language-Action and World Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RynnVLA-002: A Unified Vision-Language-Action and World Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.427156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:26721f0164ae04766a55df9ce2e6363dfb893bb266a8bd4818c3c07747617ce2

Observation 3546b866-9125-46e2-872e-6903d860429e · outbound

This paper cites GR-3 Technical Report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation GR-3 Technical Report

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.424772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:0792dace45ff598fbebc853beeb078c799c361f3dd1d9db28e8e9d55f0835e7b

Observation baae2a12-7ed8-416f-a29f-65fc208ddc12 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.429643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:95df3637bb3f1129cdff435c9925ddb81d433ff37098be7a79c28c3d7fd1b862

Observation 69b1b96c-2985-4494-996b-88bd41ff2eaf · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.379180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:ddc7c72650ad2828af6d45bdc94495327538cefa6f932467ce48d95ab079052c

Observation 00efa83e-7c76-44d8-ae8b-4b4977cf3de8 · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.381185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:13823e4dadc671234b19f2c6252bff864aed8984ff8c675b0aed6d0b2bfc8ed4

Observation 5b63e49f-108f-4948-98ab-13e3850f0b70 · outbound

This paper cites ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ThriftyDAgger: Budget-Aware Novelty and Risk Gating for Interactive Imitation Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.409160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5d51e58228cacb9e2e2b2548255cb15e855d06d1e66941a3d88b3395ea52eb2d

Observation f3531602-0d1b-40cc-8a46-a9a56b74f387 · outbound

This paper cites RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.410282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:67151c050fe7fd90179b974a4932249d7fb1fff58772710e6647fcd03b3e30f4

Observation f25ab638-847b-491d-a10f-6304913606e0 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.403716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2e7ce2ea2a234543ad454665ee37caf73344b59fea761548a5f07afc5a5526dd

Observation 6502b855-75d6-437b-b569-94d3fa4dfb64 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.398578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:299ec68c3777ddbed043feef3d21c4b1068dd632409e43a60fcf0f4ebfb7ad47

Observation 5f1ceea0-ab82-46d1-80f0-3a8b7b3fb9e9 · outbound

This paper cites Galaxea Open-World Dataset and G0 Dual-System VLA Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Galaxea Open-World Dataset and G0 Dual-System VLA Model

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.401207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:5fa9bf87e7c477d6fd62a5fabb6f8db66c1252b69140244d2678e04486954b67

Observation 9971aaf8-5eb7-4986-a1e1-84f9dcfa6e4b · outbound

This paper cites Hg-dagger: Interactive imitation learning with human experts.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Hg-dagger: Interactive imitation learning with human experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:3a74791ca7ad7d799c3bb0993ae596981c0b727a5fe789edf481f63ece9e3ad9

Observation 661eb9c4-c784-4ecd-9c89-3e375c28f63c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.426409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:6797311ef2858c261943790c89382e12eb61af8f604918c8f995056eabde34e0

Observation c6209d33-3571-4bce-b111-00e70158b444 · outbound

This paper cites Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884–19895, 2020

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:7c41ba2c61f7254a7b6426ff57557c279c63021a43395f622ebda10d575c0c0b

Observation b2c245e3-e9de-43e1-a95d-67ed23389d15 · outbound

This paper cites Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.417816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:379ec4762bc3820a39c12d9e59b1c0e499c9f83b6a7070b69bc5e114c83e337d

Observation 303ed987-7efa-4c1f-9f94-64bdbeb09045 · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gr-rl: Going dexterous and precise for long-horizon robotic manipulation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.396125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:91012d635647d8877a81e5beefc99c02b4330f3c29bfc8b0884eed6686399fae

Observation 7efd9857-db2a-4235-adf5-185c32fc30f5 · outbound

This paper cites HoloBrain-0 technical report.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation HoloBrain-0 technical report

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.374801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:83d24891ed880c5c81eea01c73c727fba8cd1e22208505e4c36e4d5cd692ff8f

Observation 0bfdb06d-9720-4581-ae5d-4454bc18913c · outbound

This paper cites Flow Matching for Generative Modeling.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Flow Matching for Generative Modeling

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.361856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:08d17debde55d743892d9b2863a0a7968875262e5755c2f46954eba935eb022f

Observation e0dcd32a-d2b5-437d-931b-7a0379f56b7d · outbound

This paper cites Being-h0.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Being-h0

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.387419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:2779e2893fce77d74ffaa0a23e4a67fe9e14d1b6a484c49fe26946ec2b8b7c85

Observation 02bc5ddc-ea70-40d3-9c9f-bd502fcc6c9d · outbound

This paper cites Human-in-the-Loop Imitation Learning using Remote Teleoperation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Human-in-the-Loop Imitation Learning using Remote Teleoperation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.370614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:a75054f91982df777d8a266f074138e23ef81bb2834455ee7627885b330d6c3e

Observation 6b0e00ea-2806-42c1-b3f1-4da682061fc8 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.383901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:bb84d8e0fd3853d4557f64acbe4a87e57220e272a5e7a17d3c63ffb058b5af18

Observation 282d1af0-5a82-4121-8dda-60230e9a3c7a · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:4babfbe66277a20b2378bc4cb14f7cb9105b8898a8739fb4dcbfda3d1392111e

Observation 30b04f0f-7b75-4878-8525-60865447498e · outbound

This paper cites Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.390427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:26cb7bcab43e00c64381d7d91118a62345efc9ef6d0918683ca670921585d5b6

Observation 144a15e2-87f3-4934-bac9-a89fd6d816c5 · outbound

This paper cites RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.411949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:f582cf25e55113d985023e6eeb88350945538e9ab3b4c4836a7b092bcbcd4d4d

Observation a827281b-00c4-4db8-a95e-0a4cb8330a04 · outbound

This paper cites A Pragmatic VLA Foundation Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation A Pragmatic VLA Foundation Model

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.415298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:b5303eb478ec72cad9900b4db37280f9fe882f6364072db527d8ddd62f9e391b

Observation 3226f6d6-6d58-4bd1-882d-ebfc51706b56 · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.393321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:faf3692f7022e89b2b431aeff07f77ed74b1aefc266ed1f902a9d5ac56e8ef88

Observation 1e35cda1-f2a5-42f0-b844-59321df3fb55 · outbound

This paper cites χ0: Resource-aware robust manipulation via taming distributional inconsistencies.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation χ0: Resource-aware robust manipulation via taming distributional inconsistencies

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:761abdeaa57ada8fa0983d8f5ca43299aea6f765dd276ace93fd40dfffd94391

Observation 5d4ce01d-0765-4e20-a9ba-43dea1aa92a6 · outbound

This paper cites Igniting vlms toward the embodied space.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Igniting vlms toward the embodied space

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.376197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:f0cb2402f301d797d37bd7b1c65e805d02b0603624244b09160ddeb7ea2610ea

Observation de67d6c5-8687-4df8-acd0-8354226005d1 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.353861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:46f8912ac95a7213d516cde4f5f05b1dfafb6fcdb6fdeb486c8b431ef97e8bb8

Observation af956799-8209-4fc5-81da-52d926799d12 · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:56:10.363522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:7604ef273221516767540ffda8b7171618708dd4d496df30d8c9316403d8df69

Observation b3ef4f82-d55c-4413-bd89-f10d377a319c · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Rt-2: Vision-language-action models transfer web knowledge to robotic control

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T21:58:48.611343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:9ebe4ba576c26216e7ba7df01d546611c39ce437ca3ae6494e5dcaa30d565197

Pith citing papers

Observation a8523749-3fbd-4682-9e55-639488fd01b0 · inbound

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects cites this paper.

SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Manipulation of Deformable Objects DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T20:46:04.923901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:46:04.923901Z digest=sha256:37605861e43ca29f8f70583cff0a4bf4198c6d592b49ea355773dd3ac7078655

Observation f8f84011-ab4c-4c8d-a5ac-eb0591ac4d79 · inbound

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation cites this paper.

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T11:21:18.588143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T11:21:18.588143Z digest=sha256:5dd429273fce0500a7e8b7482fb13318e6c93389a2cc399f81d27c1617821835

Observation 1ccb8264-a599-42c5-bf2e-e74afb398eb1 · inbound

Learning 4D Geometric Priors for Inference-Efficient World Action Models cites this paper.

Learning 4D Geometric Priors for Inference-Efficient World Action Models DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-11T14:23:57.266710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T14:23:57.266710Z digest=sha256:04c5770232b083433208767aa5de7f71796b3d774d5340e4d16c876611ad15e4