Pith. sign in

Paper Citation Record · LEDGER

OpenVLA: An Open-Source Vision-Language-Action Model

As of 6 August 2026, this Paper Citation Record lists 100 of 169 outbound references and 100 inbound Pith citation observations for arXiv:2406.09246.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.09246 v3

Coverage vector

measured 100 of 169 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T14:46:35.942338Z

measured 200 of 200 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 100 of 938 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:03:07.679371Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 169 outbound references displayed

  • verified exact41
  • verified fuzzy30
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

36
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation db7afa79-75cb-4bea-8ad6-55086dbebe21 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

OpenVLA: An Open-Source Vision-Language-Action Model Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:23:25.503833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:76a20f1538924597a65d153502399a1a03b9763385a445945077871037ff94b0

Observation 0cf51699-b079-4139-8d7a-f239c660d8f2 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.942655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:5478e932b4cf09ab87499ef2072370374be83265d6b22dd72cbd66f2d018e3df

Observation 29d9c600-96c1-4df1-9998-d94d3ddb6a75 · outbound

This paper cites Decomposing the Generalization Gap in Imitation Learning for Visual Robotic Manipulation.

OpenVLA: An Open-Source Vision-Language-Action Model Decomposing the Generalization Gap in Imitation Learning for Visual Robotic Manipulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.255474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:022b9eec29d1f5577858fcf5f585944efa5c43fa30431f06b95cce43fc666c0c

Observation 6ac01f1a-dbae-4d18-8a42-75f11db5a2cd · outbound

This paper cites Ghosh, H.

OpenVLA: An Open-Source Vision-Language-Action Model Ghosh, H

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.953364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:7dd987db6f481c3f16f0854bee8c10555b28b5cd573b67faa84c2be5ba101424

Observation d75a8661-170a-4b09-afb4-29c632c15f87 · outbound

This paper cites Walke, K.

OpenVLA: An Open-Source Vision-Language-Action Model Walke, K

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.955623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:5de57616bfcefdf36ce070aa8e3c42bb789b3c157ebd058948e302b0f93d1ea2

Observation 69f71b86-68fc-4236-957a-6ae74cb5af71 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

OpenVLA: An Open-Source Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:36:04.729348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:dc21187f25aed09efc7ed09ddec0d558b3916a7d565bb03a66b38ae25aef0e5f

Observation 7568ba38-931c-4aad-a0c2-3a2b3015bb24 · outbound

This paper cites Radford, J.

OpenVLA: An Open-Source Vision-Language-Action Model Radford, J

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.965065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:a7adfc84234c9e24559890fc7e18feeea5415735bcd8967ec9326ce1c1a80a6d

Observation b9c7c7c5-49be-4dfe-bb00-253fcee1e577 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.967033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:ab268ba380a8cddbffd016955bb2d69e0b2ae1555e727dcb93527d076d252b2a

Observation 4bffcae9-0eaa-4633-81df-4bbe38fd6865 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

OpenVLA: An Open-Source Vision-Language-Action Model Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.267867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:f9a52a34a723431bf0058d1dcad19c5e015feddafe6ea977328c839d7c044e18

Observation 8bc2094e-1075-49d3-9f81-3e9a2c14a309 · outbound

This paper cites Khazatsky, K.

OpenVLA: An Open-Source Vision-Language-Action Model Khazatsky, K

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.977435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:b288973ae07e7c0a9f1d7bf45581268970ef670672b22b023452b1f88f12d3a7

Observation 0058a515-2aad-4285-9bfd-b96f65f1f223 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.981805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:01d36097a72fefbf874eefe9353adbcf06fbfb8e13a9167caf5f13988527eda4

Observation 188bc2a8-d207-4d2d-be05-5feae508b631 · outbound

This paper cites Language-Driven Representation Learning for Robotics.

OpenVLA: An Open-Source Vision-Language-Action Model Language-Driven Representation Learning for Robotics

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.271123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:b5fe5c3987fe43329cba954ef33e971ddb9ef13bf2cedcb3c5e23448c21607da

Observation 0cdcf1a5-7c50-4148-be1f-523a9852a227 · outbound

This paper cites Shridhar, L.

OpenVLA: An Open-Source Vision-Language-Action Model Shridhar, L

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.992993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:afde9643a220e2cedb7728482f50363af5a7a1ac6376e070552c03b3cf5a3398

Observation b8513001-5337-4b36-9899-284f6eaff38c · outbound

This paper cites Open-World Object Manipulation using Pre-trained Vision-Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model Open-World Object Manipulation using Pre-trained Vision-Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.281682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4f475b98faf5fb32c86d26c73c5e413807754400698bd4a3eb65f3de1bfe45b7

Observation 9470930f-2e22-40e9-ae66-9da1e7b4f620 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

OpenVLA: An Open-Source Vision-Language-Action Model PaLM-E: An Embodied Multimodal Language Model

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:29:30.136237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:8c85d3966afd048d37b955105960393c6e5121b3f78aefa9e78fd3fd5ac842cb

Observation b2e9fa2c-28d7-4fcc-b4af-9e953ab10d8e · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.004626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:0329becc4f429309d7311c2b43b36ff4e4b3d7f2c1c40f2c8c0f57c39f7b22e4

Observation 112b7671-be40-42a7-9cba-77766597adfe · outbound

This paper cites Lingo-2: Driving with natural language.

OpenVLA: An Open-Source Vision-Language-Action Model Lingo-2: Driving with natural language

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.011402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:95d2034986da4545224657512853729959b19dff6b79c36c7f4f35fb5d2cf4d6

Observation 48818825-a06a-47d6-a88d-f79b4ccceff7 · outbound

This paper cites PaLI: A Jointly-Scaled Multilingual Language-Image Model.

OpenVLA: An Open-Source Vision-Language-Action Model PaLI: A Jointly-Scaled Multilingual Language-Image Model

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:29:06.753688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:3bf32f3b2d06058166f3f6e146b5dfe617c18d8e227bb29553fa9a07087a3308

Observation 197ddb6a-225a-41f0-a0b0-87643d848315 · outbound

This paper cites PaLI-3 Vision Language Models: Smaller, Faster, Stronger.

OpenVLA: An Open-Source Vision-Language-Action Model PaLI-3 Vision Language Models: Smaller, Faster, Stronger

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.300773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:d876ccd350a50b03d0be04a434ac46289ee3240ed73bdec2f79ee3b3b805ae1b

Observation 3cb1dbc8-1c23-437c-a90a-ee0270a81d8c · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.029613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:104a562d9605b29dbd386a1972436420615e64ebb0a95cd58c653d38396d7c6b

Observation 03b9615d-bba8-4d28-98b0-8c7f7b8511c9 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

OpenVLA: An Open-Source Vision-Language-Action Model HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:54:00.336000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:08dc6bb7f555d4e83b7c45bb62926d941ed79cd32b0812f08fce9e0eb7316a4c

Observation 80ad0449-bfbe-44cd-afb1-5c1c791702dc · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model LLaMA: Open and Efficient Foundation Language Models

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.309622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:929aa7056523e62fb01005775c048104e29bcba9bde92c096cd9e1fcd8839c91

Observation b011526a-b73b-4515-a6cd-2200478deaeb · outbound

This paper cites Mistral 7B.

OpenVLA: An Open-Source Vision-Language-Action Model Mistral 7B

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.316893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:bcba4f4a27a42785b7aa4720c7af586de2a39678c4827d0ee2b298a95258ed77

Observation 9d98c612-6b4b-46ef-80d9-e0a50e3b47bb · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

OpenVLA: An Open-Source Vision-Language-Action Model DINOv2: Learning Robust Visual Features without Supervision

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.324591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:bb0bee1efc9922cf2d74f9286fd4609c1a9a0c8cfe5491a9591a03015253f5ed

Observation 4deba42c-c090-496d-86bc-d6039da2dd94 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model LoRA: Low-Rank Adaptation of Large Language Models

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.331752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4a3d3fc693fc29c9366d6e79b11ece892cc002c4329e0a97bf52dbb8f6a08f84

Observation 0d52b70f-381e-46ee-9cf2-2c261fb6fa84 · outbound

This paper cites Dettmers, A.

OpenVLA: An Open-Source Vision-Language-Action Model Dettmers, A

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.049858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:5818ae4634b6d1aacc181f905db6f9edfe142f477905c12081202f0fe3535ed8

Observation 4bb09748-9f58-4245-b635-9d9a83234ee6 · outbound

This paper cites Goyal, T.

OpenVLA: An Open-Source Vision-Language-Action Model Goyal, T

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.051903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:7e0f67c0eadd761b0feb240ca2c4754c29974d9ac542a491a4f9edd20924b5e2

Observation 9df852a7-56e1-4332-a3a1-9541e6601835 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.053972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:757ae9f0bffaff96cd2c6653e3958ccccf94cb4140291d2677ac27704eda9111

Observation f7333548-0e35-407c-9145-749c9e48a3d8 · outbound

This paper cites Singh, V.

OpenVLA: An Open-Source Vision-Language-Action Model Singh, V

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.057660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:0664a415854cd0fab47d8112d333f1a13f45490ab91db84e65430530f069195d

Observation 99dbbacc-8992-4d2b-91ee-1b28ba4c3733 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.059945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:62be6519c23fc7d10153a4782b0e1818a50c95537d49e5ff6b7a46e9f67e289f

Observation 200208ef-86d9-4faa-bc10-825dc054b7e0 · outbound

This paper cites Kazemzadeh, V.

OpenVLA: An Open-Source Vision-Language-Action Model Kazemzadeh, V

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.061961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:1af55ed05686d6c8ba94239cfd9fbbeed00775bd0b456ff4cdfe1a775dfd6e5c

Observation 123a5e49-dec9-46f9-a315-5dfa4b307325 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.071177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:cbac6287cd0ebd299ba2ee3c85395a1cb55510b4994ca682471e39b5d8cc5f14

Observation 0f44d718-30af-4b09-8bf7-3b500e3402f4 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

OpenVLA: An Open-Source Vision-Language-Action Model Gemma: Open Models Based on Gemini Research and Technology

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:54:09.189864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:850438cf484ec557ca3fd38d8dcf75c53e11da3d1d3ac4d0a4b058ae7de7ecfc

Observation d0221559-10ba-44df-a28b-24908b6d4070 · outbound

This paper cites Textbooks Are All You Need II: phi-1.5 technical report.

OpenVLA: An Open-Source Vision-Language-Action Model Textbooks Are All You Need II: phi-1.5 technical report

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:18:01.425637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:7894870f9c16133151cf921b10b77317046e2430ab974dae46b019863f0fbbf1

Observation 2f0c8fc7-7aac-413a-8e11-525e10487eeb · outbound

This paper cites Qwen Technical Report.

OpenVLA: An Open-Source Vision-Language-Action Model Qwen Technical Report

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-10T14:46:36.355072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:00f8ee5d395bbbd4d96bbb554e93482157810a1244c9b1fda616a3248dbbfa6d

Observation 2d5d5671-c8de-4f9f-9bc2-3a03de1979ab · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.087332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:76fbde97b4d5b22b8a8ae374213ac42c8ad80193cf53e4f57983122fe38aec1e

Observation d2255e37-d313-4748-91f7-cf8199c91395 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.095968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:b99a2394d2d72e9d2cf73597d56948eb671fff47131fdaeb649c2fa03c0fd7ac

Observation 829b75fc-3030-48ad-ac28-8717fcd1d357 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

OpenVLA: An Open-Source Vision-Language-Action Model InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:13:52.361122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:45206de74a69ad75a580e0d64576965a1e11c4437c824b521a669cee9bdf4357

Observation caede679-8a21-47f2-93ba-4e174906d225 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.108205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:dc7ce1e30c3afd88fe70448aa4c9a2cc910499d1539e0a910eeadca84b369d02

Observation 040bceb1-116b-4a80-8d71-8b18759e38ea · outbound

This paper cites Laurençon, L.

OpenVLA: An Open-Source Vision-Language-Action Model Laurençon, L

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:37.139921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:34dfebd079ef8f2cce56a4f57a48c2a2bcd9717f4d263cc6f6eebd406a8593f0

Observation c52dfb7f-d379-47ad-9ef8-95f2300d1b37 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:37.148899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:692b57d427b86b1b24311446bd015c6f08df98fb656ddacca16d99587233db9a

Observation eae8f4d4-3d0b-4a22-9d7d-e77a43837fd0 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

OpenVLA: An Open-Source Vision-Language-Action Model Improved Baselines with Visual Instruction Tuning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-12T19:11:34.143871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:b697ed159ea79ceec526349982f42d0f57fc2018109e3187d09ef54793b72674

Observation 8193b89f-fd85-46f4-9e06-6cda30466c1e · outbound

This paper cites Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.251110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:87eabd39a0da0d46b25322205cc50745a7a30769a1548e79b7b2e21eec261ca1

Observation 55ba4592-6f4b-411b-bd2c-22ac4e4d1f65 · outbound

This paper cites QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation.

OpenVLA: An Open-Source Vision-Language-Action Model QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.382354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4519b2052b73c3c9d33c26f3cb08734cbf7c3aa4b815415a59aeb8ae64267685

Observation 86428803-88f8-4e28-8503-7e87d93cce1b · outbound

This paper cites Kalashnkov, J.

OpenVLA: An Open-Source Vision-Language-Action Model Kalashnkov, J

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.548155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:7644fc028a5373bcb1af712f8c4216bd7942757c2cbbf77f65f7c09d481d259c

Observation fada594e-f57d-489f-aadd-844b4c368a70 · outbound

This paper cites Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets.

OpenVLA: An Open-Source Vision-Language-Action Model Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:55:55.512010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:37712b68cb2fd8ab48ea8a2cd98f12462541bd4b546375e4ffdc9216977e62ca

Observation 063ed35f-0374-48b8-b7bb-2fc8c9e40817 · outbound

This paper cites SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World.

OpenVLA: An Open-Source Vision-Language-Action Model SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.394467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:9bd1f9d3671d180d8f48a7fea5a61b06eed4e240a181824534270fbf69ac4c38

Observation f66f9769-56ec-49a1-bb70-1eb1ab28386a · outbound

This paper cites RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking.

OpenVLA: An Open-Source Vision-Language-Action Model RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.401578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:26733a43bc9f9b8c69ce4878f12feaae4e948c136f1d5e0ebbaa259d84f9c906

Observation 1e3b884b-5e43-4c4d-9be1-7b1b324aff0b · outbound

This paper cites Pinto and A.

OpenVLA: An Open-Source Vision-Language-Action Model Pinto and A

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.576581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:6eb9c7cfc79d7c15d2eb0c99b5d8e1558935411c0e6e0c62c2794295544ad535

Observation 81e56952-cd16-437e-8ce6-bc696b072412 · outbound

This paper cites Mandlekar, Y.

OpenVLA: An Open-Source Vision-Language-Action Model Mandlekar, Y

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.583630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:24b229c3f3c4893a89a51143066c6889045dd83632c275affaa9e61c9eb63389

Observation ed0e51a3-2fe2-4e7c-ab0a-0d464dbf54bc · outbound

This paper cites Gupta, A.

OpenVLA: An Open-Source Vision-Language-Action Model Gupta, A

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.587025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:f390ea15d812574ad1d31f6f434189a32ae312cd63a2fed333a247e1d5408943

Observation d5422361-9980-40d8-acdc-22b192ababda · outbound

This paper cites Dasari, F.

OpenVLA: An Open-Source Vision-Language-Action Model Dasari, F

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.593887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4ace041639bd000bca5c816aa40cf9fc5c4dbc548a0e965f6bd7959a872b74e1

Observation dc48d38a-a229-4e1d-87ab-1b616662311a · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.598196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:464813e932d027553df5006165845a608cf9e3c5c2cd2259ac3e420c86890590

Observation 9fe7a8d8-23ee-46e5-aa72-4baf86cd41b9 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.601365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:312952416c8d52c2caff54e416dfd5f56deef84f60043c98e99362b6dfc3eace

Observation 7c2ecd22-044f-4c42-a278-39597b40ef86 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.605082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:c99357684588bfbce8c941b497d893e001ba9e919b0f6d9d61f8b00426b76ffa

Observation 80e44838-2e0f-427c-bfc6-f0599eca2f4e · outbound

This paper cites Devin, A.

OpenVLA: An Open-Source Vision-Language-Action Model Devin, A

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.608770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:2bda6e7e0f4711cc3fb12810beeceeb8c9c70465e7aa539a98d8ee438d06b6ab

Observation 13214cfc-7e96-4615-a8f1-35840404ada5 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.615554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:ad8add18a95919a0e7b2981dec6cfedd97047ba7dc0d3aaaf8781ecc869a2afb

Observation b6144feb-f724-4973-9a22-c333473f075f · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.619859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:673479fa4a737cb90611a4c76bde3752c00003b3243fa3c696f6af8bde744ea3

Observation 416edbdd-fd58-46d0-8055-8ea859feaf72 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.624059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:3f84606aec66e31970368cd3347fdb4a53e6d105eabdde2d892104de7b762f5f

Observation 30536a92-8668-48ff-a830-0f0cda5a45ad · outbound

This paper cites Learning Robot Manipulation from Cross-Morphology Demonstration.

OpenVLA: An Open-Source Vision-Language-Action Model Learning Robot Manipulation from Cross-Morphology Demonstration

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.412445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4e3b77401a69f7d08d9425cc00ce1fc29ede18fe65f7542e71c239016717594a

Observation ca68c219-c8f9-4d2e-8d2e-f5c90b8f9ceb · outbound

This paper cites Radosavovic, B.

OpenVLA: An Open-Source Vision-Language-Action Model Radosavovic, B

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.635776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:2e934e30ade941efdfcddf4b587059643994e171ac6bf0201f8e5a0ec68a9905

Observation 23b7bb77-e5b0-4a80-b2a4-861280d3ed8b · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.638222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:ed8b37ad4b355e964f49ca3a5d8af3c6b66097486949e93c86dfc65031e20cc8

Observation b97233aa-1bde-46fc-952d-2586b61de3de · outbound

This paper cites RoboCat: A Self-Improving Generalist Agent for Robotic Manipulation.

OpenVLA: An Open-Source Vision-Language-Action Model RoboCat: A Self-Improving Generalist Agent for Robotic Manipulation

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.419963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:722e7794785a55fede7075426332cdcb36242cdb86eca6d6e75a99ef9c8003d2

Observation e8386bbb-770c-46cf-8bb5-82a311ab90b0 · outbound

This paper cites ViNT: A Foundation Model for Visual Navigation.

OpenVLA: An Open-Source Vision-Language-Action Model ViNT: A Foundation Model for Visual Navigation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.428955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:876d60ee15ec33b859f1a28259f462ddbf2869c84990cbbc366fcdf23998e2c4

Observation 65641d4c-1215-48f6-af99-79210b044bdb · outbound

This paper cites Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation.

OpenVLA: An Open-Source Vision-Language-Action Model Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.432809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:57882951be4be6bc5303d534ae8720213e89bd9bd5dd34b7da229b5a8f8a7d36

Observation 209b3d6b-6ab0-4750-b6d1-95247100ae1f · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.655694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:3eecd6071e15c5e9c4f84f22dadfabdf3084d88dcd1558c055da33c151b0b644

Observation c94f036e-1c78-4419-84f9-de00eec0d1bc · outbound

This paper cites Vision-Language Models as Success Detectors.

OpenVLA: An Open-Source Vision-Language-Action Model Vision-Language Models as Success Detectors

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.438229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:233e51510c66ead5a9d0ebd42939b39d58bf4605399058dd8b633febd910ebc8

Observation 6fab0a01-9bc0-4dc6-a934-77d00b4f2f96 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.661671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:68792605b6e9a9b08b190dfaabbf22c0831860bae7bbb4aebce7cffcadbf9da2

Observation ad5146b5-7b0e-43c3-9812-c5cdfc160910 · outbound

This paper cites Grounding Classical Task Planners via Vision-Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model Grounding Classical Task Planners via Vision-Language Models

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.449759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:b45c7e73cc42e35bbe4ed3c7c28208fe94001a7e95dba54ceeb19285c345a17d

Observation bca01dc2-f5b8-4aa9-b600-07ddc29a4d85 · outbound

This paper cites Sontakke, J.

OpenVLA: An Open-Source Vision-Language-Action Model Sontakke, J

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.669913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:690b087aee6f206191481fb5fb802df55735ea713927e6b691c6593ab650b566

Observation 54842436-cd67-47a5-953b-6e8296f419e7 · outbound

This paper cites Huang, S.

OpenVLA: An Open-Source Vision-Language-Action Model Huang, S

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.678865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:5f7c15b49f5f5b2f33b6a734a106f859c25c56dfd29bf9f2a3a493a0e6bfb787

Observation 8b8992ab-39f2-4bd9-9008-25defdb0c618 · outbound

This paper cites Vision-Language Foundation Models as Effective Robot Imitators.

OpenVLA: An Open-Source Vision-Language-Action Model Vision-Language Foundation Models as Effective Robot Imitators

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:44:27.700418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:77f2809298519c503419f40ab42c26fcd9bb3a90648984008d4fc3a0cdad7fa6

Observation 2d975812-61d1-4c09-b249-95e8f1228012 · outbound

This paper cites 3D-VLA: A 3D Vision-Language-Action Generative World Model.

OpenVLA: An Open-Source Vision-Language-Action Model 3D-VLA: A 3D Vision-Language-Action Generative World Model

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:18:27.368966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:782f2198324b2acd83019f545162717dc97396280e1a71743a6d05071f4232fe

Observation b6fedf3d-3425-4c11-ba47-714b0f7506a2 · outbound

This paper cites Automatic mixed precision.

OpenVLA: An Open-Source Vision-Language-Action Model Automatic mixed precision

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.688330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:6ebe45ebaac229e995a713e52c8b15bd0af1575d2a15cebdb93006c1b59061f2

Observation c8296dc2-3c9b-4047-b42b-9acf181a3f61 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

OpenVLA: An Open-Source Vision-Language-Action Model FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:39:45.042453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:61d8986c0c96060e9e8c9383853892cc8f128b5665c0ba326232bc3ef31a3de1

Observation 0548510b-2da9-4a7f-a398-8612ed8eed49 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

OpenVLA: An Open-Source Vision-Language-Action Model PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:15:20.153254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:3d0366b278a0f41ad017869f9105ea3a6b6ee838bc9ef683f527d58596cbe970

Observation 03479416-af16-4399-b510-145a9a4d68b9 · outbound

This paper cites Dorka, C.

OpenVLA: An Open-Source Vision-Language-Action Model Dorka, C

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.696800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:f1a38bf3258e3a874717c0658f24794e8998fdf6471bb52a84c281d8135ee355

Observation cc0b0235-b87a-47bc-b866-daff8be57928 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.700107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4536206b7e7d195d70fb6c09b68a6ee6505be251145b5349ab40ccd60779b84c

Observation e6c429fa-bc48-4016-8329-e616d87ac108 · outbound

This paper cites Radford, J.

OpenVLA: An Open-Source Vision-Language-Action Model Radford, J

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.704627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:e7ca16c8b526fbc843b180dec72d4aed60ae3b412da57adeff54b023aba42d1d

Observation 21ee7424-b8b4-4f38-9b9a-9d57fdbd9de4 · outbound

This paper cites Sharma, N.

OpenVLA: An Open-Source Vision-Language-Action Model Sharma, N

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.706900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:aaf5bd73ce3b0ab622751af7d53b64e8e3170caa55bb3f1bf7bdc41da7c2651d

Observation 6826c0f3-b97c-476b-bbaf-e8b291c6863c · outbound

This paper cites LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs.

OpenVLA: An Open-Source Vision-Language-Action Model LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:21:01.130308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:735edf53534281979d8eb0001c7f22c2c226ebc5e206e38946c87538556df107

Observation f73b2f1d-3ecf-43da-8dba-ce636a2c0556 · outbound

This paper cites Sidorov, R.

OpenVLA: An Open-Source Vision-Language-Action Model Sidorov, R

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.715189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:6433dea0d21b42150625059ba2abd121d4e60dd139cf28cd8caedc48d6095da9

Observation ae9eebb4-119a-4bbc-ad5a-d91a096266a8 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.719047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:d0303f60fbfb9833e7d4e2ab3de726de3bd37f00ddfbbdd7fb6120ee25a6bdb1

Observation a29e98c1-d1ae-44f3-9716-9b3867beeb3a · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.724932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:a81abfda47ffced4607b3a723ae7a31864b839abdc811143946420cf37d57201

Observation f1e3ac38-2099-4fc0-a54e-1fd79265e8e8 · outbound

This paper cites MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training.

OpenVLA: An Open-Source Vision-Language-Action Model MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-16T04:09:36.761640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:92f5674a7c24d20eeb5baf895303abeeb97b2f99339c5556390c559359d79664

Observation 61e5d36f-238b-4510-8d74-148bc6e0df71 · outbound

This paper cites VILA: On Pre-training for Visual Language Models.

OpenVLA: An Open-Source Vision-Language-Action Model VILA: On Pre-training for Visual Language Models

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.498450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:23058bb337cecfd87339e62f12372c624132f93bb70fd662d727193f0f9b414b

Observation 36cd3c2d-a051-4b9f-9543-beefc59842ae · outbound

This paper cites Dettmers, M.

OpenVLA: An Open-Source Vision-Language-Action Model Dettmers, M

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.739653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:3a9cf91725ef56d535a7c9874b23ad3b6f939e2c53ed473afb91646bdd0a3c98

Observation 3bc65c36-865d-4cf6-abb8-1c8edf056ad9 · outbound

This paper cites Tensorrt-llm.

OpenVLA: An Open-Source Vision-Language-Action Model Tensorrt-llm

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.747214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:66510e63f4ea795ea7313a956d110c56c246e9d4fac0ae8341e6da34663907d5

Observation dceaea86-28a9-43e5-9e59-b37b60bb659d · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

OpenVLA: An Open-Source Vision-Language-Action Model Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:15:36.384383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:8282a5a824aba13ed421364ca55b4d45376c6a595c48810a37e8929468c55944

Observation 54cc30e3-f7d8-4dda-a3b5-3da17e130a9d · outbound

This paper cites Leviathan, M.

OpenVLA: An Open-Source Vision-Language-Action Model Leviathan, M

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.753848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:2ed151b5334ec23b76c1d803e12744f82aa923ed2fa82e55bcb601f346de35db

Observation 17c70509-f1da-4839-8a68-8b8468f82283 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

OpenVLA: An Open-Source Vision-Language-Action Model RT-1: Robotics Transformer for Real-World Control at Scale

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:41:14.065009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:5600d0491ec3a73ae1957633e51b5ba14861694faa4989d5dc6686b811000b06

Observation 8e0cfa67-a2dd-4207-b30e-a59efa0384e2 · outbound

This paper cites Rosete-Beas, O.

OpenVLA: An Open-Source Vision-Language-Action Model Rosete-Beas, O

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.765351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:e6b9ed612cd2a0b03d3bd5dba0ff690d535602543ee3eee0c674963ed12d9324

Observation 173ae2e0-f676-45c7-9306-c961e580afe6 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.774118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:972cc36ed3bd2e5e4cc66a8e17535fa68395b0bf64859df7115c3aebe8d1a456

Observation cffcd069-f252-4261-b673-4da27be0bab7 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.776680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:78ad5af392bfd6b162be2716edea5ae43d66035293d6bc24a2bcb0b681f6c32c

Observation f8e6c2b0-caa5-4bd4-b76e-83c8e9e28d1e · outbound

This paper cites Multi-Stage Cable Routing through Hierarchical Imitation Learning.

OpenVLA: An Open-Source Vision-Language-Action Model Multi-Stage Cable Routing through Hierarchical Imitation Learning

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.520997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:35beca77fe2408ca20d7abe02121c724992ec4a7ee14cb474fc60559620bae09

Observation a34a49d3-9303-48a2-afea-37ca99439ec8 · outbound

This paper cites RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation.

OpenVLA: An Open-Source Vision-Language-Action Model RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.530484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:4768d6b771759a14a713ea6a2cf1b8a003a7535c7f835523abae2df3a18a049b

Observation 72a40f0b-d94e-49d6-a3db-91ca6b36019b · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.787029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:f46e5e6efa88532a0feb3dbdef1102b7296d92cf07611df8cb3fbf4343264d02

Observation b31e1cfc-5673-483c-93f4-3c8f5de0257f · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 100

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.791015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:d5e470a67b7668201de04b2f262e6bbb7844543a37eed9937262c226193897e7

Observation 4086578b-cc0f-4984-a551-54b0b0aa1410 · outbound

This paper cites an unresolved cited work.

OpenVLA: An Open-Source Vision-Language-Action Model Unresolved cited work

Reference 101

Resolution
unresolved
raw_fallback, observed 2026-05-10T14:46:36.795079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:f6893765b66820a53f1743b2de079e38cdda74dbc26536633782d4516055aa21

Observation 2b1a6d6f-e2d7-44b0-968d-c92cb9b6407b · outbound

This paper cites Lynch, A.

OpenVLA: An Open-Source Vision-Language-Action Model Lynch, A

Reference 102

Resolution
verified fuzzy
raw_fallback, observed 2026-05-10T14:46:36.797541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:aba7ebe23da20cc50f4d79b2a7c11b6216a30e72b5d25cae76032e5c1bbc0005

Pith citing papers

Observation 3f406a7d-c8c8-40eb-a063-18445a566956 · inbound

A Survey on Vision-Language-Action Models for Embodied AI cites this paper.

A Survey on Vision-Language-Action Models for Embodied AI OpenVLA: An Open-Source Vision-Language-Action Model

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-24T01:25:54.395236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T01:25:10.150459Z digest=sha256:5053571f23ecfe0f26d485288303e65936d60b8e21193d37ac71b3654988504d

Observation d0a3f88f-dc44-4ec2-908e-15fb29196fa6 · inbound

Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation cites this paper.

Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T12:17:01.351949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T12:17:01.294466Z digest=sha256:b2b55d58ef3de5b3111266b48b10eb0537df0c293ef25be2617d4ac346e01692

Observation bc03956f-270b-43af-ad25-83488161bcde · inbound

GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation cites this paper.

GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-12T01:09:33.815304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T01:09:33.761708Z digest=sha256:ae93408bf098b0c4b8ec751ed7bfbad2b82f8266d3c58a223cf7ae25aa171c26

Observation 968342de-621c-46a2-9d39-07384a0383f9 · inbound

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios cites this paper.

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios OpenVLA: An Open-Source Vision-Language-Action Model

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-23T19:18:20.642115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T19:17:12.954354Z digest=sha256:1b186cac0692a7343328f206090b1198dca42179aff65ddf5152bd7e3afe08c1

Observation d841886e-8292-4ba3-b8b9-8489f6aae43e · inbound

RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation cites this paper.

RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T07:46:30.628345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T07:46:30.553633Z digest=sha256:adc3c4e688faa65ac807fa9fea1f0d0bad4ef3cce9945c48e01807d9ad1a1a9c

Observation afbec97e-2146-4bd1-bcb4-7927f90a864c · inbound

Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers cites this paper.

Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers OpenVLA: An Open-Source Vision-Language-Action Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-23T19:08:21.019878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T19:06:58.600946Z digest=sha256:12b9110d0f9ca2e745070ce7e4acab19e1423af1e945e86d37d8f78dca9e125b

Observation 0ed1f587-1829-415d-87ee-b0d6b91a5e2c · inbound

$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control cites this paper.

$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control OpenVLA: An Open-Source Vision-Language-Action Model

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:37.168344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T12:38:24.425784Z digest=sha256:57be90231416fdc1d23a5e8c9af740e9b2d1573f4515358377d2073fb840f150

Observation 21a8ddec-88b3-4905-af0f-0d39e0f187e7 · inbound

CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation cites this paper.

CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-12T07:33:25.590628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T07:33:25.188358Z digest=sha256:7b3273b5b927ea05448c5ece2cf6ddad04cb0b0ae29d7eaaf8edb4b83771c352

Observation e7f02a5a-bf86-4481-a734-3c5782f024e2 · inbound

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies cites this paper.

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-23T08:22:44.306990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T08:20:05.898025Z digest=sha256:509d84ed0ab7b7f5900aa0283a0825051bc92554de9136c61ea28fa35ede5583

Observation 8068490f-d3d8-4930-90a2-bec68d52bac5 · inbound

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies cites this paper.

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:27:22.954766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T18:27:22.760982Z digest=sha256:21c7f20c749dcde2a4062da0cf22bf6abfecad58b9e31110f61340e83c748515

Observation 042a928e-46ca-4f70-822b-68c8bad0a779 · inbound

What Matters in Building Vision-Language-Action Models for Generalist Robots cites this paper.

What Matters in Building Vision-Language-Action Models for Generalist Robots OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-17T21:37:50.861026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T21:37:50.617813Z digest=sha256:3f5c250cdae31772400108f196695e8b84189db5f062f098e8537f8b781f3ea9

Observation 3a37d60b-a269-4386-82de-9d0ba9b10294 · inbound

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations cites this paper.

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations OpenVLA: An Open-Source Vision-Language-Action Model

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T18:38:11.261735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T18:38:11.110166Z digest=sha256:3a27507bcf2810f01b470360ccabae7e1f7fb226741ee2b255f248a9e896a447

Observation 752045ee-252c-4ac9-8516-4dccea3b3dc9 · inbound

Cosmos World Foundation Model Platform for Physical AI cites this paper.

Cosmos World Foundation Model Platform for Physical AI OpenVLA: An Open-Source Vision-Language-Action Model

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-05-10T23:38:45.595747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:9707e7c1034afc8c46158f2e5fa71c68ea78b579349b5c6765f8a28e72f25856

Observation 88da5c10-2c5c-450b-adec-a4abfa1220ff · inbound

FAST: Efficient Action Tokenization for Vision-Language-Action Models cites this paper.

FAST: Efficient Action Tokenization for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:52:32.015560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T08:52:31.686474Z digest=sha256:dbbe1d9cfe29058d382b5e66a094e5d6691ad1512f3cbb4d03af6db2634eeb42

Observation cbfa3c7d-53b8-4532-81c3-e07f6216e8f7 · inbound

SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model cites this paper.

SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-12T06:12:19.954295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T06:12:19.643111Z digest=sha256:9c1b03912e6811c7bf12d8cef6ae33ca76da689b3be67b3bf395d8a5d2c4f18c

Observation d0930179-e240-4fb7-9590-621e177cd281 · inbound

Large Language Models for Multi-Robot Systems: A Survey cites this paper.

Large Language Models for Multi-Robot Systems: A Survey OpenVLA: An Open-Source Vision-Language-Action Model

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-23T04:32:32.337432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T04:32:05.138744Z digest=sha256:aac040dc648781e19d0319ce6e8ca5fdeac4178538ccda79de9819d0dd159c0f

Observation 418a98d6-c4c8-4bf1-87cc-6c240f08a799 · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-14T19:48:48.950588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:144dabfd98e126bfd4b798433e71f8012db27fd07a57c3b05fc1e5d89d776fc2

Observation 99c0666b-25f3-436a-987a-25319cdc5d3f · inbound

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models cites this paper.

Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T22:53:37.371736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T22:53:37.120692Z digest=sha256:617b0485faa46cfea530487ad1b07f403991dba3e24bfa5082be71e8f3e15fc7

Observation 59e936ed-6448-49e4-98ff-08171fb18625 · inbound

Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success cites this paper.

Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:35:32.596698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T04:35:31.914360Z digest=sha256:44448a3399a8fe24a66f5972751db525acec8e05f68336b3513aa457485a79b5

Observation 7ce638f1-12b8-468e-ac71-cb5764501101 · inbound

Unified Video Action Model cites this paper.

Unified Video Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:50:29.748845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T17:50:29.675358Z digest=sha256:81cf2239fac9221b8e40958449abf600bd8fcde0ac07db03524ba878ab51b66a

Observation a9bd48d3-4443-400a-a148-a63b1247b623 · inbound

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning cites this paper.

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-23T01:32:22.559623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T01:27:33.123243Z digest=sha256:96e4a068118088473046a442e2d7ff09b1eab39aac0e094c339382242d557458

Observation b678919d-2c20-41c1-8a23-e4203be5e326 · inbound

AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning cites this paper.

AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:06:27.213596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:06:27.136345Z digest=sha256:3b8a204cd11b54fdc4a641a6824b42fccbb87ae6395ca812a122fd7af513430d

Observation d55d4bfa-6501-4b5e-a705-dbb4c9c2b1f3 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-15T22:00:49.008387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:82078e95d818bc1a30ceb03853e36138cc95d7552d520b9cb0dee23b433c3394

Observation 7be27480-b84c-4626-8665-b1f824a6270f · inbound

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models cites this paper.

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-23T00:02:17.849713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T23:58:57.819555Z digest=sha256:4010230fc6aea8afe1e6ce76bf5898161312d398b2ed3599a5ebc877a415e6fb

Observation 5e67bf7b-9dd9-4ddc-bb22-a27dd3124858 · inbound

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots cites this paper.

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots OpenVLA: An Open-Source Vision-Language-Action Model

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-10T19:09:10.219102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:09:10.112304Z digest=sha256:beb42c5eac166a26e96a2b7691fabe1c6c0094c8a6dc6c223783979acd66759f

Observation e49d3dda-a5dd-47e6-bcd9-9d1d610871ab · inbound

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning cites this paper.

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:47:10.229628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:47:10.146795Z digest=sha256:194e6b2882ea9d3cbcb2d9cb6247bf48cb9f940da153080477633999244563f5

Observation a44e3e2e-54d5-4976-a7f0-a46b0d476789 · inbound

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models cites this paper.

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-16T05:21:45.055362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T05:21:44.903048Z digest=sha256:2eb0b6ed8c6945eada251dd9d87a20068cff093de555b852a3a136549fcab76a

Observation 4165981b-6ef7-4708-b7ba-05cb1053c4ef · inbound

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets cites this paper.

Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets OpenVLA: An Open-Source Vision-Language-Action Model

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-13T16:25:00.430282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T16:25:00.365534Z digest=sha256:65c09e2459d166fff086d560e98adb6150282230bd5406f96ef83b03ab1252a0

Observation 24e088ea-b9ca-4b63-bfa1-28a8ce078cbd · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization OpenVLA: An Open-Source Vision-Language-Action Model

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-22T18:05:00.928639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:5956ed277faca7366847eae98498f9f136deb31c4accf826f7822233a8486829

Observation 650932f2-53b1-4bb8-a3eb-3db3413083f2 · inbound

J-PARSE: Jacobian-based Projection Algorithm for Resolving Singularities Effectively in Inverse Kinematic Control of Serial Manipulators cites this paper.

J-PARSE: Jacobian-based Projection Algorithm for Resolving Singularities Effectively in Inverse Kinematic Control of Serial Manipulators OpenVLA: An Open-Source Vision-Language-Action Model

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-22T18:31:55.851956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T18:28:26.116236Z digest=sha256:992ccac665b5d63d7c25ca20f888520ba6332c4f1d938520aaf5a3e7a2471794

Observation 8630e196-1379-42f8-996c-c1a127e1f775 · inbound

GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data cites this paper.

GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data OpenVLA: An Open-Source Vision-Language-Action Model

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:55:52.167894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T20:55:52.109166Z digest=sha256:8efa414374cf987ea44179c36cc05aba2cebe983c6e01187186fafa37b7da0c3

Observation 5c6a2c5c-3b48-4f53-bc58-b6edbf668bfe · inbound

VLAs are Confined yet Capable of Generalizing to Novel Instructions cites this paper.

VLAs are Confined yet Capable of Generalizing to Novel Instructions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-22T16:46:47.449352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T16:46:05.993833Z digest=sha256:ce8b7ec51299f0c6c53ab43059bbb48e92785a8325eb90bcfd18651f42c30de5

Observation afbfd4b5-3801-4611-9780-aa4725fde2fe · inbound

DexWild: Dexterous Human Interactions for In-the-Wild Robot Policies cites this paper.

DexWild: Dexterous Human Interactions for In-the-Wild Robot Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T15:21:44.582080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T15:21:21.778285Z digest=sha256:daa8592e1938020d2f4c80cc3fb4ad6a5ea949de30ff79d671ffb7fd34ccd5cd

Observation 3e4575df-3465-4024-b0b8-5cd226f59c67 · inbound

Policy Contrastive Decoding for Robotic Foundation Models cites this paper.

Policy Contrastive Decoding for Robotic Foundation Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-22T14:11:38.482428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T14:09:48.762737Z digest=sha256:3757fc6c45ab5d2ec34fc0ebbaad2e04bd7be3a3c43491edbcdecca5026af0b5

Observation 20b8b037-ef86-41b5-a05b-99056ab48ae2 · inbound

FLARE: Robot Learning with Implicit World Modeling cites this paper.

FLARE: Robot Learning with Implicit World Modeling OpenVLA: An Open-Source Vision-Language-Action Model

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:59:08.990825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T15:59:08.846629Z digest=sha256:783cf62b5c4fb2acdf107dd1b62d33661046bf47e2d105d39fcfedb5720d13ec

Observation f5ec679f-64f3-4990-82d8-d5207121d771 · inbound

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving cites this paper.

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving OpenVLA: An Open-Source Vision-Language-Action Model

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-22T14:31:40.673885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T14:30:47.787654Z digest=sha256:0b2a74aa1067fc3a8bef73594edc5eca559fa47c17641942c15ae2f314e917fb

Observation e074ba19-1be5-41fc-8671-34b34d1ae0cd · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.438028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:18d7555938111452c0c1ef20350fc607648454f925a5dda75f6fc66902a1be48

Observation cae7ed7a-5a31-4b4d-be7f-bc699718442c · inbound

Real-Time Execution of Action Chunking Flow Policies cites this paper.

Real-Time Execution of Action Chunking Flow Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:18:51.751407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T14:18:51.613045Z digest=sha256:2b18bb6b0d3d987dd0771d8efd7169ee32adf9918d350f2b34eb955d0b2b3d37

Observation dee87fdd-3f34-4d1e-919a-a6fce3010e77 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:33:50.935341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:dfd06b948ab696bf1caecf99dec161d04359336dd28d9b15dbf860015b7ad088

Observation 3af473aa-ab83-4830-847c-e424b3611a52 · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-19T09:17:14.100488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:0ddac36aebf09061d9ad3eefc94ce42a4a13d0298b0612b2ab3eb2e5b2ab904e

Observation 88dc24c7-6d77-4093-9a38-816c681ac6bd · inbound

Block-wise Adaptive Caching for Accelerating Diffusion Policy cites this paper.

Block-wise Adaptive Caching for Accelerating Diffusion Policy OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-19T09:37:13.997104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T09:36:09.790248Z digest=sha256:6a2d26c296248db235cfb627d6d7b41fce7c5b06faee9109c70238d0fc7048b6

Observation 7d15c4e1-e281-4b51-9009-ac33ff0c234b · inbound

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning cites this paper.

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:46:44.030959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T21:46:43.955825Z digest=sha256:61573de4e6317c6e66e49d88aecd700629414ec638d309ab8a89ceab91c1ba1d

Observation fe37133c-64fe-4dcd-9ea3-4d9f5f46f169 · inbound

GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics cites this paper.

GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:30:49.362170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T00:26:28.696429Z digest=sha256:79f79a1c9407fd968563f4f75db3b82a0807b808e12ace31d4ca17b89cdb1e89

Observation b09543bd-fd28-4c31-bee8-f0061bbfde3e · inbound

Steering Your Diffusion Policy with Latent Space Reinforcement Learning cites this paper.

Steering Your Diffusion Policy with Latent Space Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-17T21:55:46.446400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T21:55:46.183007Z digest=sha256:44b7cde41ae4e666b15aff7b2628234e1100378be998480e3092c6965008e57c

Observation ae0faf98-f20f-4cf5-bcbf-3531098ac7c0 · inbound

WorldVLA: Towards Autoregressive Action World Model cites this paper.

WorldVLA: Towards Autoregressive Action World Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-11T22:57:08.177050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T22:57:07.883617Z digest=sha256:c93d2755e13d9be7996683e3c3bbec5ed6911b7ed3ebe1c4b88a26c349a21d9b

Observation 4bb4f0b6-434f-4323-973a-cb532b45933f · inbound

RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation cites this paper.

RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-19T07:22:09.311532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T07:19:07.662487Z digest=sha256:9fdead06f111273cf5312ce60f2626ccb555e903f8296400178f1a4eec143e76

Observation 3130f410-b9f2-480c-be85-96f0d04d0414 · inbound

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge cites this paper.

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge OpenVLA: An Open-Source Vision-Language-Action Model

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T15:42:41.478969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T15:42:41.363422Z digest=sha256:48446535d92927727abdcaff008172a16028af6be5e1a518b25f49f3c8366711

Observation 54ea1e18-c118-4077-98f5-40bdca25bdac · inbound

A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation cites this paper.

A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:32:56.675763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T04:32:56.397350Z digest=sha256:5970089267ab80d7b0f18b5437df0b20d60dc3363690ce2d392cd6e139a0f258

Observation f84b2d9b-4b3b-4716-acd2-ea32b5986814 · inbound

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation cites this paper.

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-19T03:52:57.561589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T03:52:18.984005Z digest=sha256:705c83ecf58af2684da244113c5fa3f01215252f224365c7bbfbe520486dfcdf

Observation d01f0582-415e-41fc-9d1e-6e6308583684 · inbound

GR-3 Technical Report cites this paper.

GR-3 Technical Report OpenVLA: An Open-Source Vision-Language-Action Model

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-17T08:04:12.663710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T08:04:12.433863Z digest=sha256:0056c287e70d9473e625c92c099489efc0aac5a12ea4628b5f8b2690944d5068

Observation c7a6312a-da03-4a12-94de-0c9abb281b67 · inbound

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning cites this paper.

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-19T03:22:01.051946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T03:18:14.655384Z digest=sha256:c44c3b1c60c83e986de5166cfd04670e563e4991455594a3fde723177dac24a9

Observation 81d7c517-70aa-4b73-b302-2d84dbdbf141 · inbound

villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models cites this paper.

villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T21:52:03.032150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T21:52:02.893886Z digest=sha256:a9bdf85f99b2ffd81955cb4bd2cc83ec8d56180dc2eee4705cf7c9e9ed2a867f

Observation d38ec4df-1d52-4a30-bbdf-9b42b17424f1 · inbound

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions cites this paper.

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T23:53:43.518630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:53:43.518630Z digest=sha256:397412be565d2940e9e87d3bc6e939768710957a563e7461b27748a5f600a626

Observation d096f913-21b5-4b24-bc32-2e3a77b556d9 · inbound

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation cites this paper.

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:28:42.033212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T21:28:41.904725Z digest=sha256:9a587437cba10b819b6a1fd5ec51c033047aaa75d3fd4921b3c8a9898e036869

Observation 82a54b05-05b0-487e-baa4-a9bbeec5812f · inbound

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing cites this paper.

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing OpenVLA: An Open-Source Vision-Language-Action Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:30.780614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:30.780614Z digest=sha256:5e85a916b95ca068e58f6eb3abe662a2658a44ddeaad72892b2931de9e0b8c9b

Observation 30600b73-715f-4fdc-a488-f939addee4f9 · inbound

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution cites this paper.

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution OpenVLA: An Open-Source Vision-Language-Action Model

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T23:06:33.791412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:06:33.791412Z digest=sha256:f90e40fb7ebf4c3767215a5482e910cdaff262af2236bf1fb51cd4437b9f0e5b

Observation 49a755f2-9aa7-463e-aa7f-1faf2ce3f2b1 · inbound

Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation cites this paper.

Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T22:50:41.646627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:50:41.646627Z digest=sha256:4fee75f57ec6f0d412dc22e718007085ceebd1b3c3ad7806fa3e3c84a681a3b8

Observation feb81f07-ebc6-4b07-a0b4-5e2740a02afa · inbound

SPARSE Data, Rich Results: Few-Shot Semi-Supervised Learning via Class-Conditioned Image Translation cites this paper.

SPARSE Data, Rich Results: Few-Shot Semi-Supervised Learning via Class-Conditioned Image Translation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T22:46:40.228132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:46:40.228132Z digest=sha256:05f0dd335241ca976a187e956ce27aa088fba17fed61ab606bf801b3b0b6b3df

Observation 1b63e7bf-a121-410b-a01a-a2b45fdd0a97 · inbound

A tutorial note on collecting simulated data for vision-language-action models cites this paper.

A tutorial note on collecting simulated data for vision-language-action models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T01:03:07.679371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:03:07.679371Z digest=sha256:25bb8c9fd3bcfca7b318d42e76f3389387c64e2f662430b3d1f73a236ba852db

Observation cbf66c8f-7dff-4bcf-af3d-5e294652bcc7 · inbound

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions cites this paper.

GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T21:59:01.879712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:59:01.879712Z digest=sha256:4b362aff0b2f706635f75abc8a59556dabc623b19ec03c969f93ad23c16b112d

Observation 8214d37d-5d1a-4ef6-853c-abe71676449d · inbound

AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies cites this paper.

AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:44:56.855979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:44:56.855979Z digest=sha256:0fd174423e471c166eeb5fee922cb9cf57737cbfd3f3bf42f8de8dee75f82a10

Observation 57551115-7b26-4f82-8a3a-cd7cef3b5095 · inbound

GBC: Generalized Behavior-Cloning Framework for Whole-Body Humanoid Imitation cites this paper.

GBC: Generalized Behavior-Cloning Framework for Whole-Body Humanoid Imitation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:48:49.193322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:48:49.193322Z digest=sha256:03c0fc10543e937e236dc31b37af675fb5b92862e56eaae8e1c750627bc86864

Observation b3e330cb-760e-4f48-9904-2c73fac37f73 · inbound

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach cites this paper.

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach OpenVLA: An Open-Source Vision-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:35:44.968216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:35:44.968216Z digest=sha256:76259a893e7003bd2d0717107bf42692e99909111e4353617c7eb87898064ab9

Observation 3b853cf1-eac7-4905-922f-a42a17d648a4 · inbound

Leveraging OS-Level Primitives for Robotic Action Management cites this paper.

Leveraging OS-Level Primitives for Robotic Action Management OpenVLA: An Open-Source Vision-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:38:41.695214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:38:41.695214Z digest=sha256:e849e12f7d4c16f4f6b05cd9e27070fa5cacadcc65989f48654af881a6df362a

Observation ce2e3329-be43-4a10-98e2-70d9ae2aab20 · inbound

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning cites this paper.

Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.751963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.751963Z digest=sha256:7a3c15886c345da0726b4bb3d22c3bc3a2ba65a7b0cf4548e7f0bf1f235018d3

Observation 915e9a5c-cf35-48f0-a0bd-1b8ec90b93fd · inbound

CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models cites this paper.

CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T19:06:28.308450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:06:28.308450Z digest=sha256:415b527c0fa672a0ca721949364f3517780691ab10f7450e781804c93622eb46

Observation 8c62b9fb-624d-439b-a7ab-af0d464e74a4 · inbound

Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation cites this paper.

Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 18

Resolution
malformed identifier
local_arxiv, observed 2026-05-18T22:06:51.415288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:60d33134bfaa7abbfb3871efd6df445185691d67f215520f4ec905c62de60b0e

Observation e5bc387a-7767-4703-a8b7-64d9c63ed8ab · inbound

MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation cites this paper.

MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:43:24.505762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T20:43:24.417901Z digest=sha256:fe3bf8b80b36a3c221e2912de8b84e152dac592e64a24cd6092542e181076125

Observation c57c35a2-60b8-4b69-a508-522b104c0dd3 · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? OpenVLA: An Open-Source Vision-Language-Action Model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:48.600967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:48.600967Z digest=sha256:7671652d7af8200b53445db4e10b8ea01fe36c46d47c6159a95e77d67627071e

Observation f11bd41e-26d7-4a8f-8b93-d4fabedd88c4 · inbound

Ego-centric Predictive Model Conditioned on Hand Trajectories cites this paper.

Ego-centric Predictive Model Conditioned on Hand Trajectories OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:29:25.447330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:29:25.447330Z digest=sha256:95f81dffe13f033100abc1cd48c087b966c52483e9afee07db302522f26c321d

Observation 1103fbc3-cec6-4223-a35e-0632c19770d3 · inbound

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification cites this paper.

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification OpenVLA: An Open-Source Vision-Language-Action Model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T14:42:32.646626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:42:32.646626Z digest=sha256:c3b177b22460ba61f6ebdb7a8e501ef044e79070c90a06c76f1d54b0d33bc612

Observation 4f946d25-a9ea-4c18-a9fb-fda8d8cddf64 · inbound

Prompt-to-Product: Generative Assembly via Bimanual Manipulation cites this paper.

Prompt-to-Product: Generative Assembly via Bimanual Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T14:39:58.543621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:39:58.543621Z digest=sha256:c2574bc2a37745d9f84b09851a08cc35e8eb3b6551798152c1be3caa520ea9f4

Observation 6b8a8d33-813f-4fb5-87c4-a8b6f74c3485 · inbound

RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation cites this paper.

RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T14:20:39.593012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:20:39.593012Z digest=sha256:7386c879ccf03f118cd3c9180d8122630e2310cbb72012cc6af51d0ba54062e7

Observation fbf7261d-e799-4805-b0cc-b4f769de9386 · inbound

Embodied AI: Emerging Risks and Opportunities for Policy Action cites this paper.

Embodied AI: Emerging Risks and Opportunities for Policy Action OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T14:37:25.218750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:37:25.218750Z digest=sha256:17f174ee81f663ea6caaaa3a2d32664751537ae95b7972be1650e6a3f4fb1242

Observation da4e8ebc-cbf2-489b-a368-29d1c68da847 · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:40.831539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:40.831539Z digest=sha256:9241cb4bb03c3e3643ea02ded145178f9fe61d5b9ffbb56f5bdd1cd7ea251c28

Observation aa991821-8572-440c-addf-491c16a11ac4 · inbound

Galaxea Open-World Dataset and G0 Dual-System VLA Model cites this paper.

Galaxea Open-World Dataset and G0 Dual-System VLA Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T13:31:09.930395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:31:09.930395Z digest=sha256:5436f9960c6c27699b22c0c0740afc2a0754a62c5d7fe633a647f751bda136b9

Observation 0dff23c8-0a76-4aeb-a336-18e97c2bc8e7 · inbound

MoTo: A Zero-shot Plug-in Interaction-aware Navigation for General Mobile Manipulation cites this paper.

MoTo: A Zero-shot Plug-in Interaction-aware Navigation for General Mobile Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T12:23:06.497905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:23:06.497905Z digest=sha256:6da739b8f323a3039fe62329713b7a7d14a523d690836ba6b1f26bf87d390301

Observation 5a1259f7-480c-4312-902c-6c94a00623ec · inbound

Constrained Decoding for Safe Robot Navigation Foundation Models cites this paper.

Constrained Decoding for Safe Robot Navigation Foundation Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-18T19:11:47.466013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T19:07:18.832187Z digest=sha256:f9dd1419a44ec53ca6224e7574c1fc76dee0cca13770b4136a1b4e31af1be6cf

Observation b01682db-1b7f-4a9f-9553-d29fc0610ba2 · inbound

ANNIE: Be Careful of Your Robots cites this paper.

ANNIE: Be Careful of Your Robots OpenVLA: An Open-Source Vision-Language-Action Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T11:00:38.160847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:00:38.160847Z digest=sha256:9454481bb6d9fdf1ae9082fd047aa2eea1a86291ad6a5e348855f9c2be3355bb

Observation a8d857f0-8d2b-45f4-a6b8-b9b67151b6f9 · inbound

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models cites this paper.

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:30:23.563766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:30:23.563766Z digest=sha256:b212a5940cd50d7da59373b4efe77a3ad3c3fe02d270f619b76eb4c788393c88

Observation 7cd4d753-902b-481b-8dd6-d8dbd936ab3a · inbound

FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies cites this paper.

FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies OpenVLA: An Open-Source Vision-Language-Action Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T05:48:46.170479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:48:46.170479Z digest=sha256:d08e1ccf4e7a6fc999a19e7300d18b83e8f966f17503e60f6650319c96f183ce

Observation 815e5433-3c03-4c9f-b791-05573eef164c · inbound

OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation cites this paper.

OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T05:26:37.150691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:26:37.150691Z digest=sha256:2f8dd27ddcfbcc139e767e8c4aead1197c6c609f6c1f7c77cabad64f1a953d78

Observation 6cc2445e-0360-4447-b44f-6171d4118c90 · inbound

Robotic Manipulation Framework Based on Semantic Keypoints for Packing Shoes of Different Sizes, Shapes, and Softness cites this paper.

Robotic Manipulation Framework Based on Semantic Keypoints for Packing Shoes of Different Sizes, Shapes, and Softness OpenVLA: An Open-Source Vision-Language-Action Model

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-05T04:36:21.032619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:36:21.032619Z digest=sha256:55b3c5ad64ebeeed57bc48c5d4b41f85be830fabb9a66925e46b5e76babf78a6

Observation 49a66a05-45a4-48d8-8c9a-f22a3697f4a9 · inbound

LLaDA-VLA: Vision Language Diffusion Action Models cites this paper.

LLaDA-VLA: Vision Language Diffusion Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:29.902536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:29.902536Z digest=sha256:15bf7c540ebb4816aa592bee9f8b4a4cb223e85cb4218b8b0480677329bcbdd1

Observation 1a249ff5-b6db-4d83-ba40-88574676a8d2 · inbound

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions cites this paper.

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:42:47.037713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:42:46.978148Z digest=sha256:90016c06a6971067ab0e508adfe482309a7ccd74a8f7d47f529c5d89a5160eed

Observation d0d46be0-d879-43a7-b272-d6aba15e3973 · inbound

RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction cites this paper.

RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction OpenVLA: An Open-Source Vision-Language-Action Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T21:32:57.877590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:32:57.877590Z digest=sha256:0f058192c9e3fcf8cc5c041ec8f7b194ccb7406c78f52d40f62dbc2f81825db2

Observation f5a7ba6e-db87-464d-8ed6-85b575ccd268 · inbound

Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation cites this paper.

Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T21:32:44.458803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:32:44.458803Z digest=sha256:3d279794daf7ff653ce2bc2af57a7281309c1f1e571edf6559fd9f6911620a49

Observation 18c5445f-7d43-44f1-bfd4-8633ccd682bd · inbound

TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models cites this paper.

TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:06.738070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:06.738070Z digest=sha256:041a44cbfe5be19926ba8fe7f38b92d9a22c4bf388883f1fcb4d2fdf45dcf55e

Observation 657a3e1f-053f-4ef7-a9f6-26e2e06b5005 · inbound

One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation cites this paper.

One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T21:31:07.692823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:31:07.692823Z digest=sha256:1ea67db11d2ac8be4d09a11722b355cbf4fb1b14691aa2a5eb3a73a8fdf3d1fe

Observation aec0dd06-a3f9-47a0-af39-5fed4c227fa8 · inbound

Attribute-based Object Grounding and Robot Grasp Detection with Spatial Reasoning cites this paper.

Attribute-based Object Grounding and Robot Grasp Detection with Spatial Reasoning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T21:18:50.418258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:18:50.418258Z digest=sha256:504ebf8adc0bdda3d064587e09290c199a31c07e94ffbb203f67812aa303f7d2

Observation da9911b0-82a2-43de-85cb-588293248fd7 · inbound

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation cites this paper.

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:10:13.358029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:10:13.358029Z digest=sha256:a5e5b620b7ca68d9139c9de23168f0fa6d9cb0e16a7c875eafddbaa02b0ef0e5

Observation fe389765-247a-485d-a60f-eea2285a053d · inbound

SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models cites this paper.

SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T19:47:08.791525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:47:08.791525Z digest=sha256:49454f300fc2b4211413ba08be197ac843c011542afcd128a57a9182b4638011

Observation d7401460-3f04-4a1b-948a-d15339d7748d · inbound

Boosting Embodied AI Agents through Perception-Generation Disaggregation and Asynchronous Pipeline Execution cites this paper.

Boosting Embodied AI Agents through Perception-Generation Disaggregation and Asynchronous Pipeline Execution OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T18:58:12.732914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:58:12.732914Z digest=sha256:e22283d0ce8f899c4e4522c1954598e6d01f673647e21695aa8d9e7584e190ac

Observation 37e67aa2-ff13-451b-a53b-d5cdc4a21267 · inbound

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning cites this paper.

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T08:02:11.299684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T08:02:11.189795Z digest=sha256:58094a4fb8a098d7945ca611a0fa7d37d33f8efc64458c681fe44ed46370d748

Observation 95aadfc7-5f90-4a04-a408-a1aa659cedcd · inbound

COMPASS: Confined-space Manipulation Planning with Active Sensing Strategy cites this paper.

COMPASS: Confined-space Manipulation Planning with Active Sensing Strategy OpenVLA: An Open-Source Vision-Language-Action Model

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-21T22:10:42.640414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T22:05:55.494910Z digest=sha256:fc4bcda6506c6a7ad695e85b410689a54364df608ab95a98bc663eac3637f3fa

Observation 8bec39a7-d188-4970-8142-f4c7d9c9e2e7 · inbound

RoboSSM: Scalable In-context Imitation Learning via State-Space Models cites this paper.

RoboSSM: Scalable In-context Imitation Learning via State-Space Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T15:34:52.225441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:34:52.225441Z digest=sha256:ac269463e3c30545cfdc7ee39c95b06e70ee45dc52d832a2a2bc957099deb68d

Observation 9aa8cb66-2de4-4d6a-88b7-f5649d031956 · inbound

Training Agents Inside of Scalable World Models cites this paper.

Training Agents Inside of Scalable World Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-15T02:05:52.599228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T02:05:52.431747Z digest=sha256:0b90f41709f45dc41935821a46bd2e1baa092cb078e7bd56da858277110a7d79

Observation 69ff4193-5636-4dd7-bb66-c515479394df · inbound

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning cites this paper.

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T12:37:55.722549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:37:55.722549Z digest=sha256:02977368594e6af392fb0499d05f358d8d260b72f8f9d036e0b4cf4c38de2f4a

Observation 908c582a-9e6f-4a24-996e-2eca6a55c68d · inbound

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization cites this paper.

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization OpenVLA: An Open-Source Vision-Language-Action Model

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-17T06:20:01.959642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T06:20:01.885711Z digest=sha256:8a2516865c9c84c8617181a0ebbdbe8a1285b3af39567a9192d4c9721f56ca63

Observation df9463cd-a3a0-4c11-b176-be86bee0ebcf · inbound

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization cites this paper.

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T11:38:53.926844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:38:53.926844Z digest=sha256:cda44bc0c35f9d40248efa0d331a51604b144df8ec8382975b3018e0d8335b1d