Pith. sign in

Paper Citation Record · LEDGER

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

As of 9 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 12 inbound Pith citation observations for arXiv:2506.16211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16211 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:51:05.564651Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:08:36.570832Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:30.514537Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53acf680-4cd6-42a3-9738-8f482421420f · outbound

This paper cites Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:04.845336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:04.845336Z digest=sha256:1618e23d0b5c44f6d00b045de47ac45278d081b7c81339385bddb5e7edcf702e

Observation 24b6b7d9-3775-4067-854b-24b403a0edea · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.413572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:04.971311Z digest=sha256:9915dee49c1228b722cd594964cb1d44a5778d30b3e182e0afad6a4d6192c3e2

Observation 3743e857-a31c-4f6b-9dbd-c97790271b1d · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.020892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.020892Z digest=sha256:4c50ffccde88847382487b572bcee828bc13e4f8ed356e1604f34089cd9f593d

Observation 10ec7206-8122-40ec-aca1-ab548f5291da · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.398694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.148613Z digest=sha256:e613b21381ec5c99ef4088bfb285f72f75558a50eff6efbb6e74ffdb6a10a326

Observation 6709822b-3d59-42e6-a973-a5af90811532 · outbound

This paper cites G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.241297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.241297Z digest=sha256:8b5120e92764be11489cc16f8cee009a33319fbb7b0a149903ee40e312151558

Observation aa23ab53-82a4-434d-8ab1-a3de0f8373d6 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.384171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.305317Z digest=sha256:06445b9b54a15b83be327d51cc1eb4addd54632095d1f64608738d819e939ea1

Observation f102da75-3367-41a3-bebf-4a1a35638799 · outbound

This paper cites SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.310332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.310332Z digest=sha256:0c4dbc9f9ddf20896228d2893fbacdde518002de45053baad1eb841729bf3fba

Observation b9edca07-ac70-4075-a75c-2eb00d38d7b1 · outbound

This paper cites Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.315137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.315137Z digest=sha256:2298312e7c602421f522fd4a758a0e3b2a8f2104fcbd2f5bcbe95620a316832d

Observation 12c2b4c4-c24b-4818-bc06-6557efe69113 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.370196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.320125Z digest=sha256:5fbf0136b1d9f039b002cf5d4bebc6111dcb93a195acfb4085d6533555c977c4

Observation fb141498-28bc-4ab8-8223-bd5e1e935b2a · outbound

This paper cites DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.324527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.324527Z digest=sha256:2a534e16ce7047b34eeb5744c0ef5288ff2938e89bb9ca4d1ac849db5ecc9759

Observation b67b89f6-dba4-4c1d-859a-a2eb8a8d0e91 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.329116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.329116Z digest=sha256:ad67d8ef0aee99e5a118f2f38cd157921f333441b96db19013c3bfb722840cb9

Observation 5434ca92-d28a-4d43-a5bf-a5ffb9011982 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.355296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.333582Z digest=sha256:a6602244f1f4559a75b36e0c786025fe4304036f511e1acf2b28f8f2f0019183

Observation 2920bfd2-698c-417d-b44f-3398e71fae21 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.340945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.337888Z digest=sha256:589e06d6d344bb9016d360bb9b10cb11cdfe5f018df818d7d1307086c65572b7

Observation 0a42fde1-05fa-4965-9ab9-194ad301421d · outbound

This paper cites Huang, S.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Huang, S

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.326178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.341678Z digest=sha256:f279bc4c4dd45604f2570ac9c6cf9874a84c797ed65cd315683916556097e1a8

Observation 44ca8117-ffa3-4ba3-8b2f-4f61a5c76260 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.311543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.345758Z digest=sha256:a4425f2dc7ac4e7f859d507d1863cb5e2f32aec5c3e8c0a209c216488324b60a

Observation e36abc5f-663f-4985-9208-3b9a1bc7c1bd · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.297310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.349416Z digest=sha256:357760224eeaff671c2ac8ef341132bbd13c43cdd5a781078cc4373bba17ee32

Observation 6a6d9936-f27f-4121-9e17-6715b23cb50f · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.353073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.353073Z digest=sha256:4966bf96f8ef9074f382458d172ba8fc4367db5bafcf38b9821c78bb62b3c012

Observation 998ae332-334f-448f-b49c-6788db8aa696 · outbound

This paper cites MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.357203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.357203Z digest=sha256:b99f0f9e4a25cea04e4a82960cc97a2004a35ff7f813cc86d004673f8e249f72

Observation ffc1e156-24ae-42ec-920f-4664b4d2668e · outbound

This paper cites RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version).

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.361677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.361677Z digest=sha256:582b8a70b9a14a0f814d961ea5ceb35ec2cb5ba49d75569db4df210e3dbfdb33

Observation ae9e3d4e-b8a2-4919-a1c5-79bbf547d64a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.365914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.365914Z digest=sha256:887e3ac6d5f00d31aedbf145f12b4eb38e90e06c77538d30cb7c06117a7a5e4e

Observation 71d98caf-432d-4063-afd0-bcbc680928d3 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.370129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.370129Z digest=sha256:2bfa968ea8f45d55fe881f9e9a24ec22cc1b6fa65dd0c554b589953c64ced208

Observation 781c4e61-c402-41b0-8f99-a5c44169bc12 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.373806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.373806Z digest=sha256:e8dfdfc7bd5355fc965794ab387ed0a819b9320468b12772509e805bfa71012b

Observation c6baf59e-1106-4231-aad2-8e8aeb2c6d45 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.377722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.377722Z digest=sha256:d84794958f906b34a3ee31a916b79fa81bdc608d65a8afc359eac595092405f2

Observation 6f9c26fd-89b7-4d39-851d-3702d09d5286 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.382153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.382153Z digest=sha256:8093fbdf22a50d6c5148019f4ff0127e2dd4542ea8d34e7293acc840d000536a

Observation 2089d7f5-93ce-4dec-9f61-c58a2b0a3e7e · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.386011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.386011Z digest=sha256:119ac7b9d6dd83d0c338c74c0aa449c32999594397295df1c77d67578522f42a

Observation e4106095-b4d8-467f-8a49-324fed201f2e · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.390144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.390144Z digest=sha256:b806335c5f885bca740194b13c553113eda8bafe9102f78a9533607c2f42cec8

Observation a7eefe94-574b-4443-8831-d84e432eb660 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.266276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.394404Z digest=sha256:a3fa4995319d7b308222133183a10f33ebe1f0bc18ace05340c4247bc0ca8274

Observation 2776f220-336c-4073-b42a-25e4523a694b · outbound

This paper cites Zhang, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zhang, A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.398711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.398711Z digest=sha256:8abdf67eb0657f7dd656bad438e9c239d9097045045e67a046abd58c796f6d9b

Observation 31347ab5-89e4-4054-9245-a85a0314e0e4 · outbound

This paper cites DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models DenseMatcher: Learning 3D Semantic Correspondence for Category-Level Manipulation from a Single Demo

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.402554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.402554Z digest=sha256:053143cd8614c3a024d93cb12dd4921e3640d039297fa0feca30f55ee72bc051

Observation 57f1110b-73d2-4ee7-99aa-ea5662ce3d31 · outbound

This paper cites Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.406504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.406504Z digest=sha256:321478d5107e8644d7b5adacb7041e8894feb7f4355756fb1f6e9204048a17a8

Observation ab0881a8-66fa-4b01-8888-c298d62784f4 · outbound

This paper cites Tyree, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Tyree, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.244679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.410625Z digest=sha256:6833efab274eded0f41b155d0ec337585b16ef6c43dfffddd5046a4275515ac3

Observation 6027aae2-becb-4268-bb2f-2747a4bc7ac8 · outbound

This paper cites Migimatsu and J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Migimatsu and J

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.414591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.414591Z digest=sha256:1d20d34fbf04392a5827670d50e4440475e89037f62dda00d10ad3e04d2ae5b5

Observation 69e1497f-f748-45fe-a3b9-700c6d65cb3c · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.222540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.418384Z digest=sha256:48beaac5960a67fc1fbb0aec9b82137d0feec9f85c9391cdb60f0fa079b00956

Observation 37e04de4-db34-4e8a-a077-eb8902a63ecb · outbound

This paper cites Devin, P.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Devin, P

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.208714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.423236Z digest=sha256:e50bbc261cf0ca2115badf953143973c1f8d3e88fde9eb1fcfba6781a1fa1caa

Observation 6644594f-06b5-4192-9696-e405ca079be9 · outbound

This paper cites Locatello, D.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Locatello, D

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.194944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.427977Z digest=sha256:e9fbd6b50d5226d5e1de4a8bac4327763c8550692dfd86092af836ff263add0a

Observation 11d0a78b-0401-4773-bed6-2bbd59e90b06 · outbound

This paper cites MONet: Unsupervised Scene Decomposition and Representation.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models MONet: Unsupervised Scene Decomposition and Representation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.432378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.432378Z digest=sha256:3de2f99ae7f86b64303f8ede19ad6e80d94c390ae46f973ac18b8ea168da3e53

Observation c012871e-e0af-44e4-895a-ac63d00c216a · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.180647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.437120Z digest=sha256:377430ac40c7fbc9413e75ff12e40251cd5cea8a317bb29314cabe095c2ddbef

Observation 7b2b6983-52c7-419e-bb41-46b3f5252e36 · outbound

This paper cites Heravi, A.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Heravi, A

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.166949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.441912Z digest=sha256:8f0da833a52d1b2e979ad46235cc3360f157bc2f133b414115881a9bfee29cfd

Observation 168b2c4f-2d25-4599-86bd-93423394dbd9 · outbound

This paper cites Zero-Shot Object-Centric Representation Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Zero-Shot Object-Centric Representation Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.446892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.446892Z digest=sha256:73a3f418d2a789773c63a9b3a05b2c0446fe2d03881a3c11e3380c6ba80ca1f9

Observation 45466b54-eaf7-4a98-8d1b-44323e663ace · outbound

This paper cites An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.452350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.452350Z digest=sha256:cc71200a8ab172b0a0657888e121dbdca3c5bfdb9ec5a15c028df2c6f5c7a4ec

Observation 5ae50293-7817-4635-aab7-310d03f9b75f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.152241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.457334Z digest=sha256:d9034757ff86942dd52d423035c3ed5b808a32f32e09e58ef4573881908d6cf9

Observation 1c96920f-812b-42e0-bed5-ede0ebce4e50 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.135773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.462541Z digest=sha256:36bde41431dffa20b1a119c00069abaeaff4ca42eb9a87e276ddcd7e66f110d2

Observation 9c1fbef4-b8ec-4ff8-882d-0f8b321c677c · outbound

This paper cites Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Stable diffusion v1.5 model card.https://huggingface.co/runwayml/ stable-diffusion-v1-5, 2022

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.122138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.467010Z digest=sha256:d58993342c09b942f932ae9f0922c06146778f433f5e207fcd8931e14ac33b7c

Observation 9d5454d4-9a24-4458-9b93-c7625e4d3ee2 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.107327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.471119Z digest=sha256:85aeaea227beca9973a1601b4a9d7c7341ff8c157331856efd4162d5ec2a884a

Observation c1a03e2e-f18d-46e2-be08-e8a6d00adf5b · outbound

This paper cites ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.475594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.475594Z digest=sha256:cb6c3d422fac196c709b1df2c1c9ffaf04f1311e6c47f3a63794152a19b2a5c7

Observation 6802722f-9f79-47f5-bd9f-89530512801f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.092961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.480005Z digest=sha256:d61ae914e1256dd982621b0133164a324eefba685436b239faa028d16559f009

Observation 08c6105a-ee10-4731-8c6c-afa6600bf20b · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.077473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.484539Z digest=sha256:76692f624745bf38660ff1734078133e4ab5c89e681761cf454f25662c97ec4f

Observation 73232784-6260-411b-a695-60a0ac9a5d2d · outbound

This paper cites Bar-Tal, H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Bar-Tal, H

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.488962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.488962Z digest=sha256:270f9afe0d53b784254f561d5e3121affc74f3976313f0a9ee0c42cf1aa6ab8d

Observation 8814db4c-fa02-4477-bad9-47d83a39da8b · outbound

This paper cites Dai, L.-H.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Dai, L.-H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:06.053967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.493365Z digest=sha256:5c0dce5f029ea2cc3c6bc082cff247178873a5c1d3198955ccad36cce84206a7

Observation 1fef668c-b1ef-4f1a-87a7-07fbe2a7171d · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:06.039106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.498683Z digest=sha256:2886bd7426308d3600c4663abbe0b9218919648e9de4f1d3016f9bf743f29eb2

Observation 094b105c-41c1-4408-b040-87c109eb108f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.503078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.503078Z digest=sha256:4b2e4446c22399a8d60d089981701950bb61ffbfbd94a863eaa9092834d1e16b

Observation 0eda8c59-061c-45c6-a4ad-bc6afc8f021f · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.508411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.508411Z digest=sha256:6b59b5ea715eb7620722861cb47971d3569d9a849df2b99e83405b8aabd29709

Observation 257b6e7f-9da2-416b-8d1d-dd1e5d1e21de · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models SAM 2: Segment Anything in Images and Videos

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.512735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.512735Z digest=sha256:812200c0d65917e81ece4b3f5c5183adfb0c47f1924a30d6467b57b84896d460

Observation d5638925-aebe-47b4-9048-97992808df66 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.516993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.516993Z digest=sha256:7a0802e7aadca5688dd93fab6f4899c81e3a962854ff8534720ae2732914d058

Observation cb770695-661f-42d0-b598-c3c1fbeca1ab · outbound

This paper cites Krizhevsky, I.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Krizhevsky, I

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.522115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.522115Z digest=sha256:652b930a2406e31fbe7e7caae5cb77ea7bbffcc47255efc75d4af001f09408bd

Observation 9f592a5a-79e5-4b5f-9741-1bc67c350e3d · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.527593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.527593Z digest=sha256:5c0e95832fe93dcf0bf0b0f96776d4e15313d3e13abc80a39238c6c154060ccf

Observation a31165a0-14fc-47c0-90b5-3d49cf2f0175 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:51:05.990154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.533167Z digest=sha256:fbb6405d668ded22a6c88de0437c77fdb45c7f90162360782628698fcfcb840b

Observation cb1f02cd-b236-4b34-b417-427e5d224f48 · outbound

This paper cites an unresolved cited work.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.538786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.538786Z digest=sha256:34fbbbaae435b30be629f30672bfca12e5c26ad5c6e3c21af3c277a8930e43f1

Observation 53e4b447-c84d-4971-a6ea-afd045c95f2a · outbound

This paper cites Khazatsky, K.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Khazatsky, K

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.543711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.543711Z digest=sha256:3644a4a86ae8aa69a9dbb92dd559192e2f7d1ad1ddf11ab1963ae655b1bc8e23

Observation a145d811-0bbc-4c71-9008-28c2e2228016 · outbound

This paper cites Radford, J.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Radford, J

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.549380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.549380Z digest=sha256:8c3171a7e1dc2b61ca9c29af4dfc57d74e65180c30f0ce43eafb0beb76ea1eee

Observation 02556ad7-f4c2-421b-99d0-8e270d441f1f · outbound

This paper cites Campos, R.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Campos, R

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.950992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.554756Z digest=sha256:43806929c3c2d5535d1b62e9a99d0b4cb7b8a6ba901a286a6cf06451c3432b9e

Observation 6177047f-aa54-44f6-aea6-ef3cb1efdf54 · outbound

This paper cites Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Meta Quest: Virtual Reality Headset.https://www.meta.com/ quest/, n.d

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.936241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.559399Z digest=sha256:1e38165a3b51deb65b70747dd93a80cc484d13e64546b941b96f26f85c65d9fe

Observation 5fdcae0a-3e61-4d20-a971-52375e32b597 · outbound

This paper cites Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024.

ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models Astribot S1: AI Robotic Partner.https://www.astribot.com/ product-en, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:51:05.920767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T23:51:05.564651Z digest=sha256:e30598d0c65d3790bb38a7d297e5be7a8ee1593d9a7efb7a6a9f1b58eedd9d33

Pith citing papers

Observation d33b1b39-c1fc-4205-be93-e74a269e4339 · inbound

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation cites this paper.

Ag2x2: Robust Agent-Agnostic Visual Representations for Zero-Shot Bimanual Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:05:28.900178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:05:28.900178Z digest=sha256:2d45999323a2c491bcc4aef9816981de8c213962c07738ff5cb4266b0a007790

Observation fff67d24-b737-4cff-8fe2-ff062d4198d9 · inbound

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation cites this paper.

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T10:43:40.991161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:43:40.991161Z digest=sha256:7588feb1d804163caf3ec5a58b19b15beda17cc37049a13c1ab18a275458e586

Observation 028fdbbe-87a4-4252-9064-713e605b5d95 · inbound

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation cites this paper.

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:25:32.979738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:22:41.611893Z digest=sha256:b344d65d7cde6692058a0c94033e219bbd77c5e68209ce54d005567d1443c9b9

Observation 8fc455d9-e6c7-44c4-8e69-2d599f1370b6 · inbound

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment cites this paper.

CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:52.618248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:27:12.286456Z digest=sha256:d744e7eb9106139a7bd8468c93e36fa523ad319b8dfcee5826d9d79626c3879d

Observation b58a1dcb-a01d-4e5f-a51f-6b8a8fc1da92 · inbound

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation cites this paper.

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:19.021480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:50:33.134927Z digest=sha256:6d93e6e0b8f8f824ebe7da590fcd13adbd4b4fc830bbc2f82c8ed6c7c6d88f3d

Observation c4febba7-8387-4ccb-bdd3-5d0ebb6ffe1f · inbound

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training cites this paper.

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:11.595737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T08:07:47.793895Z digest=sha256:d3376b8d754ce230ccb6712e7b6d21c7326ac0a1185bf410899f48d96cffba05

Observation 5813eb5c-32b6-470d-a900-9fd4888a5281 · inbound

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning cites this paper.

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:22:14.583427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T04:22:07.953637Z digest=sha256:df8e485badde6a4f81597011501506dcec898581b1e5b1a6f3808c7b421b2c98

Observation ed470334-f0ba-422c-922a-939f8f57033f · inbound

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation cites this paper.

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.455458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T06:57:41.245418Z digest=sha256:2b80ba10551db439a4216b7f24039d17a3f49420afe1f1316e118e33a8602235

Observation 27e3a5be-f2fb-4e2c-aae8-907fc7f80c20 · inbound

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models cites this paper.

MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.825517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:55:53.760595Z digest=sha256:9975d63c15042f1663189ac3c194749061bbdcf57287f818b695e7f50b78699b

Observation 19db67d5-ffa5-443a-90d6-5beb6bb7b0f7 · inbound

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining cites this paper.

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:37.882421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:23:36.941801Z digest=sha256:b0aa34206a758681e6074308af72741b6c84d95b1ceb661a1855248a3452fff8

Observation b2fab49c-65e1-4a40-94a3-3accfa9d07fb · inbound

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation cites this paper.

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:30.517227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T18:01:08.616612Z digest=sha256:e5b148ddf3c098399970cee50a94e4893758d33d56c365446f8941d1e37ba6a9

Observation 49371b32-4dc8-43cc-a7d5-2b9dd8e113be · inbound

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation cites this paper.

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T20:08:36.570832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:08:36.570832Z digest=sha256:c18aa10ed170a4d13f48c2e8ccaa6a8566a3375b3d2a6f3309e301bed5572b85