Pith. sign in

Paper Citation Record · LEDGER

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

As of 18 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 28 inbound Pith citation observations for arXiv:2508.10333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10333 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:47.949431Z

measured 89 of 89 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:27:01.402929Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 016cf2e5-6ffc-416b-b2ae-b2216e04b144 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.157327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.157327Z digest=sha256:4e36772d591c881dd7d3110db66f21b5f022467a2073c60b4e5b456a2e8938e2

Observation 51f44ab3-2844-4713-9497-581cd0c11de9 · outbound

This paper cites write newline.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.256938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.256938Z digest=sha256:5449c05c4e2b99158ad510355bfc2b16c8ac4da7b0e40a44f2a6f1a0a9655632

Observation aeeb2696-acee-489f-a136-bc444a61c5b6 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.335373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.335373Z digest=sha256:ca46bd1b6acdf8b4386c2476383bb7208ea3dc3464c445331b5fa75055176039

Observation 55b44d0f-e875-4922-be3e-02fe9f8fd682 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver PaliGemma: A versatile 3B VLM for transfer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.445023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.445023Z digest=sha256:6eec33a0e0eab45c7351152a3262c06c375cd7ca92829e1f22c3b142555cdbfe

Observation c00ad31d-fa76-43a3-90a4-ad61a16998c0 · outbound

This paper cites R.; Finn, C.; Kumar, A.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Finn, C.; Kumar, A.; and Levine, S

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:55.339761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.532847Z digest=sha256:310b46b478b43cfc89ae8dbc12cbac1e59c2a2894a72aa2338093facdf47c96f

Observation 06acae65-edc6-4903-8305-99374a8069ba · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:55.111512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.649791Z digest=sha256:ba908f8bf57c8a66b2bb88ef293664ec43122a0a52ff646505378d47e431b0cd

Observation 198b9eed-5966-4a5b-925f-906ed4ef824f · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.733209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.733209Z digest=sha256:c7d0229a8df643d78b3ecf1aa8820d876e6466545d9cd45a1e1236f9905213e1

Observation c547134b-6b44-49d3-8058-63b9bcfabe75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.891716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:42.870859Z digest=sha256:3d84626c90ec3e474b01933f1c5d38cf2d9438553e6e57407c70c3061db9f8e3

Observation a0da2fd6-ec44-4efb-b3ca-6e38b27f6f4b · outbound

This paper cites WorldVLA: Towards Autoregressive Action World Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver WorldVLA: Towards Autoregressive Action World Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:42.981102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:42.981102Z digest=sha256:3e218bec21e8d118d49bd14e7c19b24cac89c9c64efb268e8b32e919461af008

Observation ab4bafc3-28af-4703-879e-7fe899692918 · outbound

This paper cites Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.065743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.065743Z digest=sha256:970529cf4d2329ef267a444a1b895520c7a49e86e9ec61be5172926bc37a412c

Observation 714ce289-afb3-4d49-997b-99ce921d69b2 · outbound

This paper cites 2016--2019.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver 2016--2019

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:54.667834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.146408Z digest=sha256:c84b4bffcf64d7b82fe7ab6501571fdbd0cf8065a90ca0074034e40713731b8b

Observation 69e98311-e672-4d82-a735-259fddfe5f32 · outbound

This paper cites OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.242879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.242879Z digest=sha256:624da06b0abe6eed12b1e4d18b4dcd1312e2388355db8a428b1367756457ff12

Observation 551c752c-1b81-4542-9afb-cc94e77c9086 · outbound

This paper cites GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.318557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.318557Z digest=sha256:0bfb4b3915fdf37c432bf86bef6f7bb5d1b2e877dca029704839ce233ed96095

Observation 938d2317-9e60-46c1-a3f9-030549681f0e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.442791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.423867Z digest=sha256:b204d03d4eb1ee3ee2b59707103c743ff232b4b388ae8df0e2369b72103fddec

Observation d35117dd-cc1c-4143-9414-317818c7bc8b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.217460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.559992Z digest=sha256:c37534cefbd357ec724d8929f92bc67a9d766fdd0763c63eb52defc499eac01a

Observation 53a7fbd4-1b51-41ca-82a2-d25b7325feb7 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:54.003307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:43.659725Z digest=sha256:f56326d386a101531a34cbe9537c7bf44e94d610a1e3df2ee6adbaa1a2e235d3

Observation 1202615c-7b2a-4fb4-ae69-df1faa8315fa · outbound

This paper cites Prediction with Action: Visual Policy Learning via Joint Denoising Process.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Prediction with Action: Visual Policy Learning via Joint Denoising Process

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.779814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.779814Z digest=sha256:b03504df69eb73b354def4292b524dd19541cf5a329eb5884cc9857bfde31cee

Observation 5ea62d55-fde7-489e-a6e8-98527acce49e · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.866349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.866349Z digest=sha256:a5cb7ae8ff8364cd28e48ac33b1cbae1c973b2b0657a5356cb31f48ecc5f92b0

Observation 1689ff9b-2280-4c5d-98aa-43d842ff1973 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Denoising Diffusion Probabilistic Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:43.969912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:43.969912Z digest=sha256:36a6dee59299b3f183d200081083ff4662a7f3239fe8e84c23c52b6b1b8847b2

Observation 6087975a-809b-457d-9aa9-8bbee619fd25 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.809157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.077675Z digest=sha256:c18b27dd9759705e38881e96831120a03845cdfae2ec98f1e792e6d392e32167

Observation ab0ec46e-731b-4f0b-abf5-efcf72132ebd · outbound

This paper cites Elucidating the Design Space of Diffusion-Based Generative Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Elucidating the Design Space of Diffusion-Based Generative Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.168164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.168164Z digest=sha256:45b977b968ff09b447fefa7f2f9db5f7fae381b88c439c5aac3ce21d1dc796f2

Observation a9876cd6-eab2-4244-8bdd-b648e0c9944c · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver YOLOv11: An Overview of the Key Architectural Enhancements

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.246525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.246525Z digest=sha256:d8eeeaaf6f662b1c86716abb569851296d0d0bcff0bc884c0e9115b182d9a817

Observation dcc27518-8402-45ab-8d2d-e921d14404e3 · outbound

This paper cites J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver J.; Pertsch, K.; Karamcheti, S.; Xiao, T.; Balakrishna, A.; Nair, S.; Rafailov, R.; Foster, E

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:53.600727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.342040Z digest=sha256:ff268db29eb9149eaef964ad5e96f4850523dee819c136291b7a5b7c99b2fc31

Observation da5c8f71-2c0d-4347-a6e9-a711655e2601 · outbound

This paper cites Auto-Encoding Variational Bayes.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Auto-Encoding Variational Bayes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.447704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.447704Z digest=sha256:2a0f54700801baf5d4df563f2f13acfb5a333d6ed538d9b10d066e7a5dc6b44c

Observation 9fca09ee-667d-40b0-9ce7-6108f3bd8d1f · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LISA: Reasoning Segmentation via Large Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:44.554728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:44.554728Z digest=sha256:32b720abe454b459f86c2b070582a243aceb887a2c78ce51663a4354e14cff45

Observation b70c07cb-7301-41f9-bb17-c5cbbd37e5a1 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.342812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.669023Z digest=sha256:e2a2272aaf97f02373d8b5548502678038cf61bcaf08cf2c2185a576b6325d14

Observation c6d3a858-2cef-4a18-868a-14a5e7a7fa19 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:53.096576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.765284Z digest=sha256:8a0b74263f5379cf3a4de9491e360bf241cea9629c75fde21be3de99221f3d61

Observation c4ffcb0e-a213-4910-8ad6-d0482e7ffc4f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.824198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.866675Z digest=sha256:37218ae9843886ec7d9c5e99ec936a643e501c272b44f6cc2cf19fea952547f3

Observation 07a1d701-131b-4fb7-81de-748660bc9c78 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.552907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:44.956249Z digest=sha256:e987129ff95d5d1c7e2218812886a15374e9c14bfaa0c2014671a00727f36dbd

Observation ae041e0c-9b5f-4608-8038-a44070d822f3 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:52.262406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.066281Z digest=sha256:3727b0a34bf35472a7cc9f5fdb78ad04bf6309c55d5833f1d9e0b00dd384085a

Observation cc708740-53df-4841-9066-d53e6a38ec97 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.212485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.212485Z digest=sha256:986dc26de65740ec069350b1a008923de577d82306f4438300abd43b44dc9494

Observation 19505c9c-a177-4790-8423-3693c94b90c0 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.940390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.297037Z digest=sha256:05a9dc3bf88b7181a5a32c2db9f9631bbf80c66e9b9e1dc9de79a3f53327e661

Observation dc6d773c-9e4a-4678-94a5-6d73163ec34e · outbound

This paper cites LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.390488Z digest=sha256:92cad2ffc958fd10937179615139796ea3f187449a26ef06e1d20c375561c42c

Observation 89925814-2c7a-48a3-96e5-87396f0cc01b · outbound

This paper cites Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Y.; Sanketi, P.; Vuong, Q.; Xiao, T.; Sadigh, D.; Finn, C.; and Levine, S

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:51.617250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.482182Z digest=sha256:05e89cee686f6768bbe3646d9bd643a9596fad28e758d4ce277602160f806568

Observation df7fb02b-7e2b-4311-a8b1-676ab029924f · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.298656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:45.579047Z digest=sha256:31fe7c5f6443896f037f96bc78e3767edee2028710cf4aa3bcbd89a5c2c90a16

Observation 8276046b-c484-4278-8264-49318ee9d348 · outbound

This paper cites Scalable Diffusion Models with Transformers.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Scalable Diffusion Models with Transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.709820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.709820Z digest=sha256:b14265d8fafa139b728524b7f3c7413fc23e019a2a356caa3ddf75e6565fb601

Observation 1cdedee4-544a-4af8-a9b4-0b3a8e676f7c · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver High-Resolution Image Synthesis with Latent Diffusion Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.824230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.824230Z digest=sha256:0900da947132692f6de443bc85cf6eee13ba51147bfa2bd7896745c007452fd7

Observation be481c99-2867-4366-a198-bc684b3d9df4 · outbound

This paper cites CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.934582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.934582Z digest=sha256:eaa8422e539cc1874a84e1604a7cc962bf07dd5b92219fb03060739ff67ccaa1

Observation dc72166d-d190-4471-abc2-81fee668cfae · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.054694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.054694Z digest=sha256:c433d52113861c18e43854bffc93bf0b56695deaa1a1f9564645da456dad39aa

Observation aa00de08-9610-42d6-9bd4-d8f86cb31bc2 · outbound

This paper cites RationalVLA: A Rational Vision-Language-Action Model with Dual System.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver RationalVLA: A Rational Vision-Language-Action Model with Dual System

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.126998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.126998Z digest=sha256:8c5dffef2a020337bdaee5010db5e8d8bb9775e7f8a63753b5995a0cced1351e

Observation 962310b2-29a4-435c-8c94-a14bacd57d74 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:51.007699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.259758Z digest=sha256:3b2de5306c31f8252017b476d54cf2e7d512c8bc06795f8bfd8e82367aad17ef

Observation c8a4849e-a5f7-4daf-ab83-e3007ad3de02 · outbound

This paper cites Generative Modeling by Estimating Gradients of the Data Distribution.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Generative Modeling by Estimating Gradients of the Data Distribution

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.370745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.370745Z digest=sha256:89fb80ee6a4bd0a9b40c1fd2aca187d053d46036addaad61c2ade1865d7f152b

Observation 555eac16-8dc4-49c0-8c20-45642aa8b9e4 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:50.728268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.479452Z digest=sha256:1fc146ee6295f8be99975af29670182e09071713c0d074f2d077495e4ce71b7a

Observation d881a641-d32d-4186-baf1-a0d8bba7cae6 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.559044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.559044Z digest=sha256:b1b314230bd63f1ae85af3cc40ab9ee17535946848b06508c9a907b611da839f

Observation 04fd7e57-eaf3-4ceb-b882-4eea9e122eba · outbound

This paper cites QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:32:48.264716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.642262Z digest=sha256:c44766febeeb4f0a0c31c22e8586594bd529d537ce4f13631ecded6d2ecbb8fd

Observation 78b3481b-d7df-4f6c-b1ad-bf84737163a5 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.725496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.725496Z digest=sha256:6c6657b1d003434ed61e0bd0221c7fa6a8d47df74894832c55f48e99f1eafdd5

Observation c1cd9d61-8884-436f-b469-261b8210cb6e · outbound

This paper cites BridgeData V2: A Dataset for Robot Learning at Scale.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver BridgeData V2: A Dataset for Robot Learning at Scale

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.783305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.783305Z digest=sha256:a68a21590a9f77c84a44195666455b8351f0daae2ef2c19ea016b759c96e0cc3

Observation 5e261572-3b55-4c88-b4be-eb504e83cbf3 · outbound

This paper cites R.; Black, K.; Zhao, T.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver R.; Black, K.; Zhao, T

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:50.320404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.873909Z digest=sha256:257f126f7028b204a77588c2c778f2e3a43f00269064ac5f6fe9ae13a395f076

Observation 961c6b19-704e-4cf5-8b3d-43130886e1a4 · outbound

This paper cites Reconstructive Visual Instruction Tuning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Reconstructive Visual Instruction Tuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:46.962180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:46.962180Z digest=sha256:326d268a8d5364001096cc7fb2828ce6918fe853ee60064aa4dd02211cc4c3dd

Observation 35480d38-3d8d-435c-a75a-6db8a4a4b601 · outbound

This paper cites Unified Vision-Language-Action Model.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unified Vision-Language-Action Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.040368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.040368Z digest=sha256:cffedbab83024112281e36c250c2f0c79d6f6a5bb487ded0c1a744f0a96b43d0

Observation ab9de928-841d-4d9c-b2d6-9bac4f94c513 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.945212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.125046Z digest=sha256:39cab473f90e3abd99b8c6793532838744880770fa3e28152d722c5d3d07c917

Observation f7912514-9f6e-441a-a431-939058de3eed · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.736982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.234493Z digest=sha256:49a8d59a8be853479f4cbb836701b9d72cd3b8ea4b9f210058072b4e5119e390

Observation e3675e46-39e6-40c1-8a83-8b3963868738 · outbound

This paper cites Qwen2 Technical Report.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Qwen2 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.323607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.323607Z digest=sha256:a28e941921887d1199e6583bbad4aa3a02f93fce16656fd1be16fd561b79b904

Observation 9b121ff5-6903-468a-b8b1-5072aa742e75 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.535472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.410676Z digest=sha256:3e8a73fcd2da4e1557d6d21d5eb7bf482d5ace145dd6d3682c62a366e3f1028b

Observation c72d3213-3f3d-4590-99c1-76b9ef2cd74e · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.498001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.498001Z digest=sha256:45a73e8cd50707b63cb05a30d28dd60ccdb8177ed01c480c3ea7bf14080a5ec9

Observation e5c2536f-6e29-4fd2-9a13-780548522aeb · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Sigmoid Loss for Language Image Pre-Training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.563746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.563746Z digest=sha256:e65ac0a2ec5a1f451ec649c5ff8f2e4a36effae22c72401cbe117f6218180c02

Observation 4e7d3ff6-6072-4b93-ba59-62fe787ce832 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.340025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.644076Z digest=sha256:71094970ff5f9fc2ca1c99d7c2a04c27260285ad62d2ee899e64ed46e45ebcd5

Observation 2878e17c-26e5-4437-902a-464de73e7722 · outbound

This paper cites MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:47.702884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:47.702884Z digest=sha256:647ef6ce80c52008b7aed756950ddc9e23bc3a9f7dd22c36ea97d95983687869

Observation 30b05cf0-1631-4f7f-802e-65e17540558d · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:49.158036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.766165Z digest=sha256:f5ab7b763086ac7b97472f88854967b73acebc40fdb891e1fd41a68dfbb2f6c0

Observation c5a20b2a-d22f-4280-bcaa-1753625b240b · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.985244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.881596Z digest=sha256:15fe20bb756131a81ecc17cf3c1dabc7bd59c32e1eeebf41a3e000f56aa2afdf

Observation c1bca469-ffdb-400d-bdd3-e767492d0800 · outbound

This paper cites an unresolved cited work.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:32:48.824489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T20:32:47.949431Z digest=sha256:a0ef15685d5f7d714d6570d3f5c92015919e53a939f74bf5bb74fe3637ca385d

Pith citing papers

Observation 19af9271-5ab4-4013-9b05-dd7b6cf08d83 · inbound

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models cites this paper.

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T09:34:44.154129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:34:44.154129Z digest=sha256:f16da175896d411c9bda13399d36f87376180c807cd8537c0f0f00e46af6f3b4

Observation 0fe86c0c-8d69-4c55-ab72-8922135f7214 · inbound

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention cites this paper.

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:29:09.956059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T06:28:22.652509Z digest=sha256:f78da842d20d71779665a61e6bf70f49ffe2a4805bf121eabc246c010498a807

Observation fa147f79-109b-48ed-8d65-e62d9182f7f1 · inbound

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation cites this paper.

Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:24:11.084402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T13:22:16.242427Z digest=sha256:e55b5f40fb154548fda4fa89be475e8084a81783496ae435d8bf407bb1c1488d

Observation 687336c7-bd92-4c6d-9459-ecc1d17ac7e6 · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.270525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T09:49:26.446868Z digest=sha256:d1180c462d1c6f6ee7ec143aab42134329b2244e4e83817819f9c73d42ca5f94

Observation ac8c78d4-7a12-45ed-a724-9bff414a72fd · inbound

Towards Generalizable Robotic Manipulation in Dynamic Environments cites this paper.

Towards Generalizable Robotic Manipulation in Dynamic Environments ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T00:13:51.580603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T00:13:51.580603Z digest=sha256:96f7227cdaa1506d637fad8eb2b14ec4b065a26fc648b32b0776dad4c090366d

Observation 61734d00-9f8d-434e-97c7-60a46713027d · inbound

Grounded World Model for Semantically Generalizable Planning cites this paper.

Grounded World Model for Semantically Generalizable Planning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:05:32.233313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T15:05:29.465402Z digest=sha256:8748e8511de4a8eb86e36abfc99c8fa511f3d1a7a9f09c7cbc61db066a766f9b

Observation 37ad6c86-c8ef-4dee-9050-fe0b8a291e96 · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:05:29.539890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T14:01:50.429917Z digest=sha256:d60091d5a5109c7e11e8944dbe46c21d827b207fabf2776b7190ae1e49aef7e9

Observation a89019a4-ca2b-49da-b1b5-e39e3a5f0e8f · inbound

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System cites this paper.

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:52.410964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T01:51:46.849464Z digest=sha256:4d5f479d39c1eb07482cf19bc8f47ef629c57cc6a4de710230408cfb1d39584f

Observation 8b519540-cf1a-4cb5-bf3d-dd72203cd8ee · inbound

Mask World Model: Predicting What Matters for Robust Robot Policy Learning cites this paper.

Mask World Model: Predicting What Matters for Robust Robot Policy Learning ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:06.115031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T02:14:17.676675Z digest=sha256:1733131a0de7fcf46b906170478961fcbedc8bc59bb1e9b4694a8e893b2c167d

Observation 8596f8aa-4994-4f4a-bdb5-6a505effefb7 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:25.118343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-09T22:07:24.208555Z digest=sha256:3f1b779839a39d0db09aab0198249e0fcdc986dbf8610bb55940172f5f86f0dc

Observation 0b5e01cd-11ba-4333-a186-a3d1e3e36d45 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T18:39:49.173828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:39:49.173828Z digest=sha256:3e1c4d217564cbec918a022ed35e5e061e9b3edde570c011d94920a8343c1849

Observation b834c9a3-a4f5-42f6-9030-286defb1a253 · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:16.014701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T03:13:36.437080Z digest=sha256:24a767dacafdf9e883d834780cbee251781f3550c7207bede160914e4651f180

Observation 120946ea-c5c6-4db0-bf1b-2c4431ab940b · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T18:17:43.564400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:17:43.564400Z digest=sha256:068e5e76d85b18d849981b1ea2c426eb260252416ad05481f9de1d6184b6451d

Observation ebe3b632-0602-4730-bbed-32727529d609 · inbound

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models cites this paper.

CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:26.986030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T04:07:52.566360Z digest=sha256:71d1d7a0eedcacc41685a2e59157a8e99051a68a01cb34e66288b76c725a1536

Observation 633e7192-f355-4783-9050-9e01dab00802 · inbound

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete cites this paper.

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:57:17.450518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T04:55:23.145492Z digest=sha256:a91af21aa4a7a4f0a113b1bed5afdc6cf2bc448657c95926471d285636e8a2b4

Observation 3f82cfcf-3285-4442-82fa-fb2cff8805a0 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:06:08.676554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T06:05:52.461348Z digest=sha256:ae893fa59fe919ccf3dcdf8ce9b3d5328ad32f2512a3ab510dc4bb83ab6a2b79

Observation b45b1ea2-afc0-4140-a249-687f7ee2fdc2 · inbound

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model cites this paper.

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:58.253115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T16:52:05.565650Z digest=sha256:1dc6f89e726a8a5881f4186281aa2cdf8d32f849a89dccd80bc8c8e4a0cd0aed

Observation 9fb6d44f-2a71-4c89-bed3-bc563272dbff · inbound

GEM: Generative Supervision Helps Embodied Intelligence cites this paper.

GEM: Generative Supervision Helps Embodied Intelligence ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:43:28.942277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T13:38:27.263726Z digest=sha256:1198a698624b91cf168dcd6646332385a18de70e3d443cf62adc5dcfa4a661cf

Observation 97f062d7-ed85-48d7-9bdd-387dd4e394a0 · inbound

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving cites this paper.

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:02:45.777855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T22:57:01.742536Z digest=sha256:af3485be870e41d8edf321e3be06c123106a920513da91471ad89e89ddc6deae

Observation c7e9f867-668a-4f8d-abba-10e39c14a3f0 · inbound

OneVLA: A Unified Framework for Embodied Tasks cites this paper.

OneVLA: A Unified Framework for Embodied Tasks ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:26:14.110096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T17:05:08.124096Z digest=sha256:b3a0c1194848b72a7172de72ec43aac1d8885e29d8472dbfa84d60956cdcdcc2

Observation def8b06d-2585-490c-9a1c-43e1c96c8825 · inbound

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding cites this paper.

AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:16:59.252395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T01:23:02.576098Z digest=sha256:9f62463c2f2c3080baaebd7599ea86bdadc9623bb2c795db8b20674d87a17c7a

Observation 679555cc-f140-41bf-872f-17c2adabb100 · inbound

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging cites this paper.

TS-Mask VLA: 2D Temporal-Spatial Masking for Vision-Language-Action Model with Effective Bridging ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T15:19:27.489381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:19:27.489381Z digest=sha256:0c38f8189c53142b6167e272ffc71f7d7e7e9f862cee83db2cf695cbef216438

Observation 432ac8cd-3b5a-4d82-bd7f-38db2c5679f7 · inbound

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment cites this paper.

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T05:16:41.306473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:16:41.306473Z digest=sha256:e31d0aa583b3e2c19e8704b776fa345f7b4a48520c809ee5d9ab9373a4d63f07

Observation 7c3528a9-9a6d-477e-9437-b8c2a6af53d9 · inbound

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation cites this paper.

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:30:08.940833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:30:08.940833Z digest=sha256:ec3c66e6433377ac7ccd77b6afd9cee44bde8f577c844c6fd5c5832fd5bd04dc

Observation a229022d-8beb-4610-9978-6d1de4977e9e · inbound

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight cites this paper.

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:43:02.573394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:43:02.573394Z digest=sha256:eb073428515c73fe6b53515663afc3d496c295d7739de4d7ad52d214960b0435

Observation f2086ba7-6ef6-43ed-bffe-b0771fb280dd · inbound

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight cites this paper.

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T18:04:05.199011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:04:05.199011Z digest=sha256:ed52dac0f47aa509a2d3106f937cfec2ab8dac8821beddd87c37e0029864950e

Observation 1041b97c-be14-4232-a8e7-8e454e373539 · inbound

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models cites this paper.

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T04:29:48.065577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T04:29:48.065577Z digest=sha256:80c51a053f0e07bc68b73115c46708c431e07791b5da11b98c83e7b3c06c323c

Observation 6a50ebe8-ed3b-405d-baff-8aefda09b651 · inbound

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models cites this paper.

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-14T04:27:01.402929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:27:01.402929Z digest=sha256:712a423dc69fc14c1c3afcce36622a77f1543af3b05843358cb0cb492c2526cb