Pith. sign in

Paper Citation Record · LEDGER

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

As of 14 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 43 inbound Pith citation observations for arXiv:2505.23705.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23705 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:45:25.866011Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T10:19:06.294730Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:07:48.099929Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 76a80ee3-69cb-403b-b1ae-61e625f80d5a · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.226085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.226085Z digest=sha256:0c7f92769e085c33e293b085028558d8b3120c0f36bd2d2370765bc829677214

Observation 30eedd93-7000-4caa-bc10-be4e7cd20566 · outbound

This paper cites Minivla: A better vla with a smaller footprint, 2024.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Minivla: A better vla with a smaller footprint, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.438033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.438033Z digest=sha256:f59e72f9af805e9659ade19485655af0a734a0fa71ae596dbdf216eec73dfe2d

Observation 78993390-e199-4423-a0a9-f91e69f4c3d6 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.481541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.481541Z digest=sha256:09d5e720bd9975bf3d407b17452199c124bcb13bc8478488dcc2b664f31bfa67

Observation aabf2d38-860d-4d1c-b40a-9be0612bd63f · outbound

This paper cites Roboagent: Generalization and efficiency in robot manipulation via semantic augmentations and action chunking.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Roboagent: Generalization and efficiency in robot manipulation via semantic augmentations and action chunking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.551903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.551903Z digest=sha256:ff0ddb7642843fa7476df22e3318b63c689486eaddf2be8193118343653d0c1d

Observation 137b3f9f-0de5-437f-ba80-2393d69cc5cc · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.621870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.621870Z digest=sha256:0078c471d7085e249109861fe019535c0f842e791faaf7920d320ea0b2d1ba11

Observation 4ebd4971-00f4-4ce3-aca0-6ffad8f203f8 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.694442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.694442Z digest=sha256:8982b6e1b8ed45c423064e547de29d17b38523e1160bd81028868e0ab860bdc8

Observation 03836f76-469d-48ec-99f8-2784393ed205 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.765111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.765111Z digest=sha256:12512e260da78133843800b57b5bb7f99af52be67e7f2eb1d7b949f97f7121ac

Observation 838a416d-83b3-4a8b-b34d-ddd73a390609 · outbound

This paper cites Visualgpt: Data-efficient adaptation of pretrained language models for image captioning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Visualgpt: Data-efficient adaptation of pretrained language models for image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:29.295864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:21.848000Z digest=sha256:a712088d1ff9c4ddb6d49ff2f2f08f36f2d5dd93074b8aba306b5390f97dd05d

Observation 1c05503c-49d9-4772-8aaa-d1d79aeb5f1a · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:21.939925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:21.939925Z digest=sha256:25ac5c46168c4777081ea83a6779622a470e632e4aa699bcf53a3d81ebd3e714

Observation 9c91362a-176c-418b-a555-dcc08a669fed · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Diffusion policy: Visuomotor policy learning via action diffusion

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.021786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.021786Z digest=sha256:71cb23b3b6bb1baec9ca1e8b14b5d6af4dc9e46b744c335fbf7cdf41d252f1d2

Observation f783929a-97a5-4608-b146-fc2e468610d5 · outbound

This paper cites Robonet: Large-scale multi-robot learning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Robonet: Large-scale multi-robot learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.062774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.062774Z digest=sha256:9cdc6bf41bd13c5eb8b806e200306a2da5e7a4d23d9fa2f95407d1e7653d9cab

Observation a42addfa-6c58-4b2e-b520-00d1023ac1a0 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.126578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.126578Z digest=sha256:1b2d2c5b9fd3794f9269c8709d73240928258b1e915ac669f09e7507ced94482

Observation d3a306d1-2a88-4af3-a8ea-06764e821cf1 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better PaLM-E: An Embodied Multimodal Language Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.177470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.177470Z digest=sha256:3a1db12f5df78f536da7aa71fe7ae0eca19c45fbed7a664abb3ba64cb8f8b295

Observation 81083a88-f5c3-4ea3-8a62-f55bf8bf6bcb · outbound

This paper cites Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.228807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.228807Z digest=sha256:ec2ee6b475066b24b02abf5e614eb5ca25baa50566f972e018b7794b396e2b42

Observation 805109b9-70eb-4395-9dec-66e8398196fa · outbound

This paper cites Scaling rectified flow transform- ers for high-resolution image synthesis.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Scaling rectified flow transform- ers for high-resolution image synthesis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.277119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.277119Z digest=sha256:1aa402cffc3298f8093bff7f93f51f3bfea2dc814e329fedb0270ed433ae0f03

Observation 75a14eff-87c4-4860-9271-e589ee7ddc81 · outbound

This paper cites Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Rh20t: A comprehensive robotic dataset for learning diverse skills in one-shot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:29.052182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:22.331710Z digest=sha256:964ef44092aa420a890cac9ddb3df75d71b421183e23af6b7c9c215495fd9474

Observation 2a8dc569-df73-4227-9319-e751988a9e6c · outbound

This paper cites A new algorithm for data compression.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better A new algorithm for data compression

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.407838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.407838Z digest=sha256:11e9f54874a781cb8e53e7d1ddd06036f7f02d514067a930ace5508835bcb40c

Observation a46d8ce7-0886-4816-8e2b-48be8317b62d · outbound

This paper cites Making the V in VQA matter: Elevating the role of image understanding in visual question answering.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Making the V in VQA matter: Elevating the role of image understanding in visual question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.455066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.455066Z digest=sha256:fbf519feb35a7e9194064eda00f89ff1ee176b21d6f123d7e93a320dd91d6971

Observation 5f6b1350-4891-4d6c-9600-4cb58b413264 · outbound

This paper cites BAKU: An Efficient Transformer for Multi-Task Policy Learning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better BAKU: An Efficient Transformer for Multi-Task Policy Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.506973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.506973Z digest=sha256:20c855954356c616b03cd2a95cae6845e19dcb2f3c851a710ae497a39ef9f7d1

Observation 1a9c0649-ac14-4158-aea4-641a5347d251 · outbound

This paper cites Otter: A vision-language-action model with text-aware visual feature extraction.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Otter: A vision-language-action model with text-aware visual feature extraction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.566886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.566886Z digest=sha256:276b5fb97803d16b7cb1f761ec16af3349a137feb035970cdcc6249583ac85dd

Observation 0f1b08bc-917e-4f7e-abd0-33b07b426c47 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.619559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.619559Z digest=sha256:20f8963b1a131db2845af344d2e014e678f78edad6f8b1b4524443e7f97f5bb1

Observation 5bb36aa9-0b1b-4aea-88d6-edff6dd68b79 · outbound

This paper cites an unresolved cited work.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:45:28.895705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:22.669666Z digest=sha256:4b687061614a620da05ce28d22e5aab7e96848b5a62da08bfd998fbf6096d593

Observation 77bac6bd-f424-41f2-9a30-db5a45a75ec3 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better OpenVLA: An Open-Source Vision-Language-Action Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.723077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.723077Z digest=sha256:c408310d9842f56878a14e06900cdca4dc61937cad57007387ad44a7353333af

Observation 55537301-d75d-4f0e-8b4a-4b9db57dc2de · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.795560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.795560Z digest=sha256:d874f231af3cf2d48a0c025edccf2db5e481833b1fbfab90df7abf8f7e23db70

Observation 790af437-ab24-4d8f-bf2f-b73bbc29e7c0 · outbound

This paper cites Rush, Douwe Kiela, Matthieu Cord, and Victor Sanh.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Rush, Douwe Kiela, Matthieu Cord, and Victor Sanh

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:28.658113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:22.837040Z digest=sha256:5129b403af77e3b30f0afc63a92d502b797651d7f45b2f246ca3c173e941ed63

Observation e00d3bc9-b00d-44bf-ab31-8918b4021106 · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.889864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.889864Z digest=sha256:c989c4feddd3dcb924a7e8438854189f1b77a2ad1237fafe6a15ff2949c9233a

Observation ef4992e2-b93e-4bec-a281-c4b8a6a40a6a · outbound

This paper cites Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.943587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.943587Z digest=sha256:6edc639e92552553fac497186bb9602ad01c74223bbd00cc7841bf12bda4e8c1

Observation a1ebaeb7-aae2-4b17-a2b8-5b2aec736c8c · outbound

This paper cites Flow Matching for Generative Modeling.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Flow Matching for Generative Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:22.986141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:22.986141Z digest=sha256:d2187c8a955559df8d7c49f8f5fc0a9ce6a178cd6c60aa5425eb7ab2f80106e1

Observation 82fd9928-6b84-46f8-86e4-17cf081081df · outbound

This paper cites Libero: Benchmarking knowledge transfer for lifelong robot learning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Libero: Benchmarking knowledge transfer for lifelong robot learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:28.421977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:23.062123Z digest=sha256:d803afb12d114ea369b39a22095020a6a93f19ab40728dd11bed6489511dec69

Observation dee705af-e1b2-4880-a3f4-a8bd84ddf87d · outbound

This paper cites Visual instruction tuning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Visual instruction tuning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.141366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.141366Z digest=sha256:9d562b3962a5626a37958a801e9c92ee99cc13d29baa9d95e418de24a2ab1026

Observation c0cd81d9-b1fc-4136-9e9b-76f31fc501a6 · outbound

This paper cites HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.218135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.218135Z digest=sha256:3f6ef5a02be5fcc81a2d1c529e6f34e26a6edeef8acb32ed7793bf9e149a97ce

Observation 01b19051-d06c-46f6-9423-65d924c1c97f · outbound

This paper cites Rectified Flow: A Marginal Preserving Approach to Optimal Transport.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Rectified Flow: A Marginal Preserving Approach to Optimal Transport

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.313322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.313322Z digest=sha256:b23c49df6d99b2f46876eb2723e71974c14ee70fa0b84882ba6a520f7783889e

Observation 37ce83e3-a302-494a-8da9-14360fff7bf9 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.411939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.411939Z digest=sha256:16474784e730280e9549d87d9e67608b849c2e8c6e53753971bb378f2aa37587

Observation 34304110-6c8e-49ad-89f3-6278434c6ed4 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.496333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.496333Z digest=sha256:90cde3c7a4c01c47a7096e9a5770dd1f466f0d051ef4b5d57bb58577de429273

Observation 7c2c7512-6277-46f8-802f-7d2bc43cf8fe · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.546515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.546515Z digest=sha256:6ea94894e0ec2a073c0cc9abf71ae21842542282f73ae8ede7c45c82af6ff47f

Observation da545aab-c77a-4f60-829b-50b5620710f3 · outbound

This paper cites FAST: Efficient action tokenization for vision-language-action models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better FAST: Efficient action tokenization for vision-language-action models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:28.161430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:23.595019Z digest=sha256:d5ee90c97c70c91ac5306bfd351194f7636152bfb51846544a009d7bf1cec67c

Observation 5fb243e5-1d95-4546-b2fb-4aa324e0c279 · outbound

This paper cites Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.666609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.666609Z digest=sha256:5c6b019054a4b918ec33cc70d93839169abe7ffd5500cf73c429656f0187a5ad

Observation 9341900b-657e-4e59-8c9a-108a24aceb67 · outbound

This paper cites On Bringing Robots Home.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better On Bringing Robots Home

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.757810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.757810Z digest=sha256:d92184614155b573816b1ededad683cb9dcfde16db6c2262a2236234f81e1e9c

Observation 31f5f6e2-7a61-446a-af72-fef9c92bf846 · outbound

This paper cites LMFusion: Adapting Pretrained Language Models for Multimodal Generation.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better LMFusion: Adapting Pretrained Language Models for Multimodal Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.846395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.846395Z digest=sha256:48a0d50c7502b8dee7c7d25bd040ea2d36347581814994be055ed6c9a5cb5d74

Observation 42cb4d42-bbdd-4218-9e25-dc0bd74d1b1d · outbound

This paper cites From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:23.965626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:23.965626Z digest=sha256:2d351015d2902d0f0ec812555e354270e50cfa462c4f82d0c5c3e3147e5c83b4

Observation aaf236d8-31a7-4d89-9239-532e89b4a234 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.134368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.134368Z digest=sha256:2d7d9acfb4e739e08ac4bd6c8852d063b0604cebc33e3d55f370e5528db39331

Observation 4b23d730-033a-4bf6-abe8-02bcb680371f · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Gemini Robotics: Bringing AI into the Physical World

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.273203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.273203Z digest=sha256:5006c0bec51b0fd13ffab4afa4776fb643b701b28a31b2b8c8ba2f039a748b53

Observation 51a1e0e8-1725-4cdd-a1e8-7dde95006766 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Gemma: Open Models Based on Gemini Research and Technology

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.334768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.334768Z digest=sha256:24d35880395e89f07cfcd6d47a4ec6c69d1b19bb863583423a3c3b627381df2c

Observation 6ff853ab-88f5-43e4-8aad-2a45d3197ee5 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Octo: An Open-Source Generalist Robot Policy

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.415117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.415117Z digest=sha256:40cc6fb5c6ccc487be7f737f235a50c7d3a0c5c8aceef82d2593b42710af3c42

Observation 1295c1f4-c026-4be5-97d5-34768e4818c9 · outbound

This paper cites 14 Cambrian-1: A fully open, vision-centric exploration of multimodal llms.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better 14 Cambrian-1: A fully open, vision-centric exploration of multimodal llms

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.924298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:24.481852Z digest=sha256:db24df3a17435acd63de4bbc22532771fcb60dbd430b6397000399b9e67d5a33

Observation 0f228ed4-3113-469a-9f86-2cc31c234fd9 · outbound

This paper cites Attention is all you need.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Attention is all you need

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.563742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.563742Z digest=sha256:3ba084922c214e1694000fab99dd118a886f05486982d39c1b0d73e23cc725d0

Observation 6a0c71c5-287e-4719-a5fc-364435d1d308 · outbound

This paper cites BridgeData v2: A dataset for robot learning at scale.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better BridgeData v2: A dataset for robot learning at scale

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.755947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:24.616346Z digest=sha256:9a1bf299a547aadb7d05176b35354e8cd06eedb7ba100237b39a55af6c207438

Observation 19d53a50-b97b-4aca-af42-b6995bf6e986 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better CogVLM: Visual Expert for Pretrained Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.669871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.669871Z digest=sha256:67b3633e4b1279e99936a2800baf930971f841f8892ef3aeb3c7abe85d2ca934

Observation c39ac880-293a-4637-9154-8e7eaffe139d · outbound

This paper cites TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.737934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.737934Z digest=sha256:c74fdceff6f0489b7ead042b4c629bae915e3baf64d12918aacd131621c7bdb6

Observation e7743152-135a-4425-8db5-bfeea86294c4 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:24.831267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:24.831267Z digest=sha256:42347fc4db67d51640610a13b2c81c82362641a2496def082e3ec5b84ee86f71

Observation 1bacdf38-2ac1-4172-9db7-aa5e09b2ea98 · outbound

This paper cites Capsfusion: Rethinking image-text data at scale.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Capsfusion: Rethinking image-text data at scale

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.629559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:24.930112Z digest=sha256:916132cbe545a61594d891b93ca35d722d4b85aee32cff27e61f7cfaca107ae7

Observation e6dc5f02-5dc3-4122-aa53-93f1e99e3c03 · outbound

This paper cites Robotic control via embodied chain-of-thought reasoning.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Robotic control via embodied chain-of-thought reasoning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.487525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:25.048202Z digest=sha256:b8c87d3daaf2534b435f9668791e56560f6c991149f32243f1a4161217904dd1

Observation 4c1c9706-b81b-44a8-a2ad-b22b3442bfd8 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:25.188985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:25.188985Z digest=sha256:bfcd39c902122852c41b24dee00cba8097e40c2d12236531a24d726a658d6302

Observation 0e317b5e-de6d-413b-b37f-6116fd36d8ba · outbound

This paper cites ALOHA Unleashed: A Simple Recipe for Robot Dexterity.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better ALOHA Unleashed: A Simple Recipe for Robot Dexterity

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:25.335266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:25.335266Z digest=sha256:95d38f423cedceb297b31fc47120c0c3532d6df0cb13fb9f28510681e092023f

Observation ae070d1c-4485-46b8-ab14-35feabada817 · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:25.445765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:25.445765Z digest=sha256:41dad6ee7a5e9232e95751d0e2b40e4c4b32a35147d301b17567f5059c785b17

Observation 733c8feb-edb1-4d46-8ac9-53d9364ffddb · outbound

This paper cites Prise: Learning temporal action abstractions as a sequence compression problem, 2024.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Prise: Learning temporal action abstractions as a sequence compression problem, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.381854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:25.544478Z digest=sha256:a5cb92141c75c5669bce9bd99a0d93ed16e643aa4ecf6e01d4bb24e3d11fa4e9

Observation 9ef14544-7a28-4e34-af14-1bf7e25f9159 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:25.661047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:25.661047Z digest=sha256:f734627f8b59c6dd064dd0b68530e9d3e65c40c4a82ef777cee6a1e0d338c091

Observation 98517e66-1878-47e2-9e7c-735daffcd654 · outbound

This paper cites table bussing.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better table bussing

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.274457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:25.746683Z digest=sha256:04a6fadff78d6236131116d61a3bf1b2f67c39b5c2717375aa0321efb451fead

Observation c7c4b5e5-768a-4542-beaa-50cce9d6393f · outbound

This paper cites text state.

Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better text state

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:45:27.163881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:45:25.866011Z digest=sha256:4b70c4e6b556c3ea46ea31678c851f1f2b83c7e5a2c76939a7c6358c5682df6e

Pith citing papers

Observation b8b988ff-16c2-4c2b-81a9-b914a0bf0585 · inbound

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning cites this paper.

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:34.138568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:56:34.138568Z digest=sha256:fcb5a73c008693ce9ed7cb25029fea38530e59f68094d5c1090d1377a713c903

Observation 760ddb39-0911-495c-a5fb-cd2eb2b15381 · inbound

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning cites this paper.

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:34:18.707687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:34:18.707687Z digest=sha256:4765f8e35e73e1ebd851aeb18524d708147646ef5fa7a5b64e68210c5517ec8e

Observation 5e34a647-82ae-437f-b5a5-cd6ddeefc6d8 · inbound

A Survey on Vision-Language-Action Models: An Action Tokenization Perspective cites this paper.

A Survey on Vision-Language-Action Models: An Action Tokenization Perspective Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-17T14:08:35.432762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T14:08:34.893876Z digest=sha256:d2b0128bc529de32baa4a6b6f07ac7af66ce8bfacafcbc61bfd403631ca36dad

Observation e7faac3b-4f50-4aa1-a58a-a66a155ba84e · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:28:16.382337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:a3a62a5680c33156756b71c4a69fb4018eefc831e98a0103282957c13b17f19e

Observation 52298ac8-cafc-42f2-988c-e2c81ad9b3e8 · inbound

LHM-Humanoid: Long-Horizon Human Motion Control for Continuous Object Transport in Cluttered Scenes cites this paper.

LHM-Humanoid: Long-Horizon Human Motion Control for Continuous Object Transport in Cluttered Scenes Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T17:12:57.807795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:12:57.807795Z digest=sha256:4f58e6dd18a5c54b865ed6f6ab9af6e5663c828e4eceed10731c6d9e197646fc

Observation 528efbc8-50ca-4651-bc85-2b26553861bc · inbound

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification cites this paper.

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:32.602928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:42:32.602928Z digest=sha256:9856ea5425edb95b8335bf29bd8de6d7c71709066e325398cbf9b6d23980e226

Observation 85d00bb9-6c31-410c-ac77-34618f2a9a71 · inbound

FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies cites this paper.

FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T05:48:48.110864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:48:48.110864Z digest=sha256:8d4e4068fc9928d3b19ce586a44d5f4f4eab60c54864f009b031475281fac902

Observation 279dc0ba-c25c-47be-9435-1f5c0f46b60f · inbound

Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue cites this paper.

Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T16:17:25.993819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:17:25.993819Z digest=sha256:cbb160115e7f37d98172a44d0687c552375c8ba4e3c7b94eda1d72d26f9d4b44

Observation 0b27409a-ab77-4523-ae8d-7674f4300620 · inbound

Contrastive Representation Regularization for Vision-Language-Action Models cites this paper.

Contrastive Representation Regularization for Vision-Language-Action Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T12:55:07.537700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:55:07.537700Z digest=sha256:bdfabe50487672b982d2eab69feb511a30639cc934f65c918972976d7b590d82

Observation ae648400-a51e-4c03-b66a-54b282db249c · inbound

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy cites this paper.

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:09:39.769732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T20:09:39.677347Z digest=sha256:d6f8479bdadfcfd79095c6c5b660fa054bd938ae87dcd9ac19a5543df836b91e

Observation 58bc9aa5-0240-48f8-81a9-ef529ca50417 · inbound

Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail cites this paper.

Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:35:13.286346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T02:35:13.126171Z digest=sha256:71241f74d60ac77a39bdf5a8de8acb13ccf997ec1003029176aa571bcd552871

Observation 5f83a05b-9e58-4255-9c48-9c0bf93865ec · inbound

mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs cites this paper.

mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:41:00.290350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T10:41:00.142543Z digest=sha256:4fb20ec9815344e73c489dbe66ca4acd8e98bdfc484ecace420dc302a002c993

Observation 674c52a9-1604-4973-83de-122fe380e049 · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:48.806914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:48.806914Z digest=sha256:4af877e53077400659419685a3e532e20d2b5cd88809a40ea1252fd74a4d9fa2

Observation 48fef077-ff6a-472b-80dd-9ad7a8654926 · inbound

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models cites this paper.

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T12:30:33.768043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:30:33.768043Z digest=sha256:8cc1d7b5ac5ae5027b4209937ef2873fe0079fb7bf809de1f448fe1eced896ef

Observation 7f687e60-ae2f-45cd-a680-28bc770188ba · inbound

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control cites this paper.

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:06:42.842453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T22:05:39.797848Z digest=sha256:571a69a8c2d3b73099b8df66fd5418c552f8f5e2a137128ec42412005d330472

Observation 4b1e701b-a8d1-4592-8a3d-b9a4f2d8d801 · inbound

Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation? cites this paper.

Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation? Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.295783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T20:10:54.362107Z digest=sha256:5da3315875388fdb271ca6ceb53cdcc5a4f0f18bca351c503e7cec5861774ce6

Observation e6164553-3547-4724-bb80-385e3de0fccc · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T14:05:29.586192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T14:01:50.429917Z digest=sha256:e9ba9310e2e77e5c87bbf40364c71ee3b019f9d14360727e7d990e7de88fc792

Observation 54c2f270-77ab-4245-b3b9-033ac3f3873f · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:46:52.282722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T01:51:46.849464Z digest=sha256:e8d66fb8b1c80adcf71124c7122a7a9a99c170cc191fc78c9021ed1f3b4f94b0

Observation 530440a0-99a1-4b0b-af6c-7ec5fd8279b1 · inbound

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment cites this paper.

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:54:48.452010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T00:52:05.348704Z digest=sha256:a4d719012830c17bb875db1e473d4ac91bc3c4a5ec4852b03b3e0f9adff98052

Observation c029dfaf-8ba5-4e8b-b711-b5f556409f9e · inbound

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation cites this paper.

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:26:11.658127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T02:51:27.662262Z digest=sha256:d7988691b5d64e07dfd6e5c2ddc32c14f678850c1ae52ef679f122d287787a34

Observation 4f098277-7c36-4d6a-8580-ed1cc719181c · inbound

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation cites this paper.

Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:21:29.225882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T11:16:58.104663Z digest=sha256:40728d9372e64cf2abcce87168c9720f9be4fc09cb0b493f5c4205aa0feb7c7c

Observation e555cab7-d2c4-41fe-b9fa-e3ea75361273 · inbound

Borrowed Geometry: Cross-Distribution Head-Importance Fingerprints of Frozen Pretrained Gemma 4 31B cites this paper.

Borrowed Geometry: Cross-Distribution Head-Importance Fingerprints of Frozen Pretrained Gemma 4 31B Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T00:29:17.308306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T00:26:22.118037Z digest=sha256:c0f602f1309d3978141e62f42d9515703e289575bf91045dac9ae22a4c7ad5b5

Observation 2ca9ddb6-f5d1-470e-b51a-4e875bdff141 · inbound

MolmoAct2: Action Reasoning Models for Real-world Deployment cites this paper.

MolmoAct2: Action Reasoning Models for Real-world Deployment Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:06:05.237542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T17:53:44.901684Z digest=sha256:0248d49a9c64b73898c92f2367fe7c269e37a20129d489cb1f374f7fe0e324ff

Observation 8dbdd4d7-1fdc-4c73-b31d-af0e838aaf27 · inbound

MolmoAct2: Action Reasoning Models for Real-world Deployment cites this paper.

MolmoAct2: Action Reasoning Models for Real-world Deployment Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:55:57.430704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-11T00:59:54.787472Z digest=sha256:74623fcf5fe4e771e72ed44bfd22ba342dbb569ab0cb49db2a316d2c4de0ca1a

Observation b3a01e14-1b1e-44fa-897d-25503ef0032c · inbound

ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation cites this paper.

ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:06:06.335091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T16:40:19.057979Z digest=sha256:57000a2ac167ba440ca5e8c69b783dc51994e72157117165ac1896c3ec20282a

Observation 5f790df7-b247-4afa-8dcf-7ec8807b853e · inbound

PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models cites this paper.

PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:11:26.056687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T03:37:53.732205Z digest=sha256:4f6566e366940404e851f8fa85ac7bc10221979c192eb02ac52572f4750c90b2

Observation 82611645-1924-428a-87a6-c06ab4e5f2c3 · inbound

UAM: A Dual-Stream Perspective on Forgetting in VLA Training cites this paper.

UAM: A Dual-Stream Perspective on Forgetting in VLA Training Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:28:55.091119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T19:24:57.339949Z digest=sha256:ef4bda5619dd47e0b8ffb9b52175e51dbbfa2c6e77d357eeb62605ca177330f0

Observation 5d88af64-d9a9-47d1-bbf9-71e99e2bfc7a · inbound

QuoVLA: Quotient Space for Vision-Language-Action Models cites this paper.

QuoVLA: Quotient Space for Vision-Language-Action Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:34:39.430767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T12:09:12.124995Z digest=sha256:9f8e13e9e78c904b40d776ae6578de07ba60423b5132f3003f0d0b8280c12097

Observation 2f61d6a4-3a0e-4a16-91a1-3f14590e88c9 · inbound

Rethinking VLM Representation for VLA Initialization cites this paper.

Rethinking VLM Representation for VLA Initialization Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:24:00.106559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T22:21:29.733181Z digest=sha256:379aca4124db02d62035962d0c9a23fd388360763d1f0745a1557bdf7942fb00

Observation ac8ef37a-5616-4f89-a3f3-647130ecc193 · inbound

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation cites this paper.

ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.460726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T06:57:41.245418Z digest=sha256:aec680acdd712289c569480a9e8b7442722d9e09975ebaeb3b29a5d0a8a40621

Observation 82cdda85-26c4-412f-a470-8ef21c463c5a · inbound

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation cites this paper.

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:46:33.307035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T09:40:04.685274Z digest=sha256:768a6c473f54f19cec1c574de1d46a97170b719be0bb7b43cbcd377fdd214ca6

Observation a05d9af2-a934-414e-999b-babc534a4e73 · inbound

Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models cites this paper.

Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:55.801115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T01:08:59.969040Z digest=sha256:e7a2c21079b08b5fd66c3423394cf768df2d29ed5eca5d6a4ff93e275650b1fb

Observation e9bade1d-2f7e-469b-b7a2-c1434f9aa461 · inbound

Memory Retrieval in Visuomotor Policies for Long-Horizon Robot Control cites this paper.

Memory Retrieval in Visuomotor Policies for Long-Horizon Robot Control Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:09:59.578420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-25T23:51:39.882029Z digest=sha256:ba451087d27b5f83330ebc971c89a6884a530e972ccb6afb64b2f73682fabe93

Observation 555fcfd9-d6bf-4e65-b3e1-99e88d20b9b8 · inbound

Scalable Behavior Cloning with Open Data, Training, and Evaluation cites this paper.

Scalable Behavior Cloning with Open Data, Training, and Evaluation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:19:53.829300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T04:15:34.832025Z digest=sha256:3e4e3b11d156f848f908d699dfa6c28e4448a4ccaa303e207995200d47446079

Observation 8de63ae5-689a-457a-acd6-7a1be4f21944 · inbound

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision cites this paper.

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T15:44:48.927040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T05:10:51.005004Z digest=sha256:f6c779c47ed457f7860661b25f44526a6296bbd2e46990bc6de3d859a65a6a35

Observation 7f097f73-ea6a-47a3-bcae-98e7c34d30f6 · inbound

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision cites this paper.

Training Vision-Language-Action Models with Dense Embodied Chain-of-Thought Supervision Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:37:22.346209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-02T20:29:13.282030Z digest=sha256:bb60cdde4402f99ce7ed5bf08a15ded02b15b5400e574691af0cd931f2663fdb

Observation 57fd7cbd-ad7a-4efc-a8bb-9ce611958875 · inbound

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon cites this paper.

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T12:28:07.331442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-03T12:20:11.540651Z digest=sha256:7b3a659d68b4ad38ad39e9631517dd875c86f54474f59d12063131c8fe59fb9c

Observation 7ef35232-7b0b-4959-9609-ffd90978c3d2 · inbound

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review cites this paper.

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-07-10T23:07:48.116559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T23:01:01.563768Z digest=sha256:35b00a2e4103cd3a4014b948c3ed36f96f6b705f4a91258d9e95ed72bc46b512

Observation 1c316376-2f4f-4eb9-ab3b-2832375c2efb · inbound

Artificial Foveated Perception for Mitigating Shortcut Learning in Robotic Foundation Models cites this paper.

Artificial Foveated Perception for Mitigating Shortcut Learning in Robotic Foundation Models Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-14T10:12:06.757775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:12:06.757775Z digest=sha256:415b390837ebdd241100ec0aba018bacec13cfb830d8233de0e9382e23163804

Observation cbf9c5b8-a28c-4adb-be9f-cd70fd74a5bd · inbound

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment cites this paper.

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T05:16:36.686137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:16:36.686137Z digest=sha256:f3aee8a005843296789914eddb538800865893a896a93090c5196b94fdfeab1b

Observation 5035b157-6e7f-4926-964a-47c72579f263 · inbound

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation cites this paper.

CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T21:05:52.597297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T21:05:52.597297Z digest=sha256:ae50fa60d999cd6967e067b11fc82709ba5e9ee8a521c72483946de8edda1186

Observation 39400551-17af-4ba7-9d22-07501ad615d7 · inbound

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability cites this paper.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.942639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.942639Z digest=sha256:fafcd42a6cd75ca70f51962f3f875499e5f087ecacf6bea6d03df7acf08ef25b

Observation 097e2180-09b3-429b-ac8f-ee62b0817744 · inbound

Decoding Task Progress from VLA Representations cites this paper.

Decoding Task Progress from VLA Representations Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T10:19:06.294730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T10:19:06.294730Z digest=sha256:581106eb45a2b5a937181423b87410fe04042d6a6a3a281f0847e81c5c1f4972