Pith. sign in

Paper Citation Record · LEDGER

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

As of 13 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 74 inbound Pith citation observations for arXiv:2505.18719.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18719 v1

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T12:55:40.245908Z

measured 167 of 167 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 74 of 74 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:20:19.953071Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T23:07:47.932038Z

Reference resolution

93 of 93 outbound references displayed

  • verified exact21
  • verified fuzzy31
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch40

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e8897ecc-c9fe-4353-b0c2-e1ca2f27f787 · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 28955–28971 (2022) 2, 3.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 28955–28971 (2022) 2, 3

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.568547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:f387f443ec587278939cbec74d5c06dc6a580bd8f9e510b0038b30bf8e5e9acf

Observation 56b70fc8-9dba-4266-a090-169ef29283f6 · outbound

This paper cites Solving Rubik's Cube with a Robot Hand.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Solving Rubik's Cube with a Robot Hand

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.301555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:4a70d77a7762903534b20a769bd67ea8a1b0e1470ee68979844f8c73d18873a3

Observation 65cd68cd-285e-4905-b6dc-0168a44eaf42 · outbound

This paper cites UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.416599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:b35ab24a5f22041f3b9474130afd7b4d226a0013be93a25a2fba4899f220b151

Observation 2f726f0a-cc06-4494-83a9-41e35d7c33c7 · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 24639– 24654 (2022) 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 24639– 24654 (2022) 2

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.514953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:8480ec18494e137f3d95538b6eb149b37dc445cd4953e3a771dba3b394f63d63

Observation 000b23d9-71fb-41b0-aa58-17f0b8ba2a7b · outbound

This paper cites In: Proceedings of International Conference on Machine Learning (ICML).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of International Conference on Machine Learning (ICML)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.517264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:d77ed064de9e348644138d88d140cd0e5c682e898f4b4f027fb050cffe20e9d2

Observation f080fb3e-2d6b-4701-833f-e05abeecf988 · outbound

This paper cites RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.445323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:bf70dd3bbdb02a6a3c1c2cfcb60ecdacf18f24c0ce3fd7964ba15e68688a4754

Observation 9c09501f-ec09-4c2e-98d7-99374eb8f6ef · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.480291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:32bfde1518f79d5ed71762bbd283b2b03639f165a16145ccf2bb8fffdc797ac3

Observation 9ec798c0-283b-47df-b3c8-0e7e80bd1c3b · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.502751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:283b3536ca3db8dd925c0e70a963907308c6139f8aa20b2f625c9e66a6656655

Observation a4e64950-9ae9-4a71-9e69-3c85ea6ff1da · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RT-1: Robotics Transformer for Real-World Control at Scale

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.293360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9f723631014b7e60afd597f0140fc6847967433563ac29921bbab76c49ebb904

Observation 1be0b055-e4c2-41ca-adeb-951803cab502 · outbound

This paper cites Robotics: Science and Systems (RSS) (2019) 1.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Robotics: Science and Systems (RSS) (2019) 1

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.519543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:eb8d104b68e0e7d971c0ceeb76719387994a3d8c3b8e886e2792e93f728c4558

Observation 546bf604-9055-4f27-8156-51f3e9f9c62e · outbound

This paper cites Decision Transformer: Reinforcement Learning via Sequence Modeling.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Decision Transformer: Reinforcement Learning via Sequence Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T15:11:11.482497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:a3b752c5205bb061af8ddd0d688016c8b1cd49bb5329a66c681efe2f625bd0f0

Observation 5363b06a-3865-41b1-89ea-7cb6aee83657 · outbound

This paper cites In: Robotics: Science and Systems (RSS) (2023) 6, 8.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Robotics: Science and Systems (RSS) (2023) 6, 8

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.521746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:1de79f3c086c52f0326f3a22354b7690d031109d5d5cd9fb0507947a9ce420ef

Observation 93cd170a-7594-482b-8d1a-2c81f8397aff · outbound

This paper cites Deep reinforcement learning from human preferences.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Deep reinforcement learning from human preferences

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.427495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:55edbd5a04621862ed551c4f4a61be17ec6d514435d06edf355476cba552ab10

Observation 2f947a66-9d89-4f63-8508-05c68e9a5cdb · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.434915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:06cf3d95489124f21d1470ea5b7409ebf52da93be898e2f6f185cad4356fc184

Observation 9e7467a3-69f1-4132-922a-7909a5721cc9 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.441396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9315fa9539dc0907260a8710ecc7e4bf221275e9441e133ba3674e7e3a05a459

Observation 92c24dfe-4132-4d8f-a1d8-ea6fe01d1f10 · outbound

This paper cites Conference on Robot Learning (CoRL) (2019) 1, 2 10.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Conference on Robot Learning (CoRL) (2019) 1, 2 10

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.524136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:7de4f38d676d4201347e1f31574007258d9dfd2d4ca1f4a6de173447f7e0f9c4

Observation 2a8d5f44-a3b2-476a-a574-038b0f04644e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.457595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:2d520bd3a4c60b09b1e336a331f8f5799f00d8421476857bd45166b75a479caf

Observation 9b21c88b-5944-48af-854a-7c81a2b98ef3 · outbound

This paper cites Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.461024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:d478d0798eb3e8d3270ebac18182574eb58f7b070bfb24c1a5658dcc2f92a255

Observation 068f2bd5-b33c-431c-b3e7-629093b2e17c · outbound

This paper cites SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.474285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:2f9e9f1fc29f6d81ad0d0dcf029dae65005c52dc6f15afb0fd3f11b75a6d94ed

Observation ba63a9c8-4805-48e9-8c53-c3d903a0566f · outbound

This paper cites Towards Generalist Robots: Learning Paradigms for Scalable Skill Acquisition@ CoRL2023 3, 5 (2023) 1, 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Towards Generalist Robots: Learning Paradigms for Scalable Skill Acquisition@ CoRL2023 3, 5 (2023) 1, 2

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.526505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:56dd76d8cdeb4194b20fe9a42ff65c40c0b9856d3f89ab5f6aba6dac1cc9f618

Observation cb18c020-6a25-4f94-adda-3ad3544938d2 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Reinforced Self-Training (ReST) for Language Modeling

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.483624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:31de1e35e862d2875b0fef8a769f1550e8fc925fb47858eeb17510b83349318b

Observation 7108ed8c-d69a-4c54-b1fd-20506952726b · outbound

This paper cites Improving Vision-Language-Action Model with Online Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.499761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:2cf2423a4fcbe639f8d0b9493d8c6454a8dc3182766acd2af6c4d556d749451a

Observation ae54093c-6909-4779-8fac-a860755b5a7d · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 31 (2018) 1.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 31 (2018) 1

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.528816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:a268a05b123b6cd11b6b6955b02f5e0bfd60718cd877cc2e93d937b85505ced6

Observation 1499b096-973f-44ef-8e18-4e3ffc4627d9 · outbound

This paper cites Conference on Robot Learning (CoRL) (2019) 2, 3.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Conference on Robot Learning (CoRL) (2019) 2, 3

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.531194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ee9588eac09781c7832183e68c9f678f292871c62fcff09a96aa0bf116361084

Observation 106e4186-f838-4e66-b32c-d72c37862c1f · outbound

This paper cites In: Proceedings of AAAI Conference on Artificial Intelligence (AAAI) (2018) 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of AAAI Conference on Artificial Intelligence (AAAI) (2018) 2

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.533587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:09e8cfb2e562c2afc0e9a7934e41646f6cd5cacf78cab83d51055267b49ce0a2

Observation 3824aa91-1593-4173-8e66-23ded22612f4 · outbound

This paper cites Proceedings of International Conference on Learning Representations (ICLR) 1(2), 3 (2022) 5.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of International Conference on Learning Representations (ICLR) 1(2), 3 (2022) 5

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.536376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9f8b0bbef208776a793f7890a7e0535796c84b84beb966a8d430f0ad34dce7bd

Observation 339b6e10-176d-4ec9-9ae2-d97fa84c9ee7 · outbound

This paper cites Imitation Bootstrapped Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Imitation Bootstrapped Reinforcement Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.322580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9a5b4ff665865eeb2d54fb9446bbaec35cb4a041e641b907d9354961161ab239

Observation ddc4c65d-dd76-4906-8d39-cd67b10ed9b8 · outbound

This paper cites FLaRe: Achieving Masterful and Adaptive Robot Policies with Large-Scale Reinforcement Learning Fine-Tuning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning FLaRe: Achieving Masterful and Adaptive Robot Policies with Large-Scale Reinforcement Learning Fine-Tuning

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.331583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:a01401c734979e07a60000b28b3ff48778751dcf633c71f775aa1c9ea0b555b5

Observation 61e0752e-fcb9-4129-bfa0-6a5bc4c7d35a · outbound

This paper cites Causal Policy Gradient for Whole-Body Mobile Manipulation.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Causal Policy Gradient for Whole-Body Mobile Manipulation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.380217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9c3e5d0736f949df1409e4473b57037d0b50dab5b904770188fecf6562f08b4c

Observation 2e95b3cc-a89f-4026-a3f7-b425f06e79ad · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.387001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:c9b6741f4cc5f71910f1485df8826677215b04d89d1b5b3a373c4285ebcf9f47

Observation ac38b858-7daa-4825-a3ab-1555b7dfe533 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.409260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ff4f8b7bb5311ce1757a4e2aace68ef58e31bc31569a836b678a402f0fd6796e

Observation 45587987-2e58-4c9b-ba7f-001bd17d22eb · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.412498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:3cc6317369695af347e32f6a97cd372a9781681f10055965e33c5948e1571f64

Observation 2602e46e-fc63-4d1c-b56b-703d9047ff93 · outbound

This paper cites In: Conference on Robot Learning (CoRL).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Conference on Robot Learning (CoRL)

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.539654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:f50cf8e983a0494953db9d254799c64963bc447b41fcae2abefc3ccb3348787b

Observation 64091f06-2072-49fe-9128-d1612f70ffca · outbound

This paper cites Never Stop Learning: The Effectiveness of Fine-Tuning in Robotic Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Never Stop Learning: The Effectiveness of Fine-Tuning in Robotic Reinforcement Learning

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.420276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:29c6532e51c42a5a6923e18a58d013c3cc808041e29fa7496254778b99758bc4

Observation 897fa41e-adc0-437e-9c31-68dfd8d45e93 · outbound

This paper cites QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.423602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:99e3dbd5f1d6112f2d451a6b729e9f84e2b893266c1035c9ebefb11fa0592244

Observation 67becbbb-e5c2-446a-b944-a51c50f7f95d · outbound

This paper cites arXiv (2021) 1, 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning arXiv (2021) 1, 2

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.542750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:1051882dca9ce765e16b0635d6016b62aac9d821176f8662bb20d96893ec5a76

Observation 5de4596b-0f31-4d6a-8b65-aed54bfc2bbf · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.431541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:863209d0e4f292150e51bc6152ca9fbcfbda0859fa9df18c17bb5c234166f199

Observation da7d9ac1-268c-4c8c-b53c-ebd767ad9931 · outbound

This paper cites Journal of Artificial Intelligence Research75, 1401–1476 (2022) 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Journal of Artificial Intelligence Research75, 1401–1476 (2022) 2

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.546268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:d7a8dbc2dd6e04ec2f271f6378a116902c50ef4b8247a57b7b7bddfe5b4daba5

Observation e074ba19-1be5-41fc-8671-34b34d1ae0cd · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.438028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:0e638bc7797059d60b7add1b86406542e94248c0d743af40d62dfde66d34c1e5

Observation 366f85c5-4aa2-4d73-883f-526c0384cc89 · outbound

This paper cites In: From motor learning to interaction learning in robots, pp.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: From motor learning to interaction learning in robots, pp

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.550002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:740bf186d83f73362257dc7be1d77cb9375021759fcd9d41b03ec0eac3bdda09

Observation 41d6a677-848d-4126-bad2-2904e7b2bd4c · outbound

This paper cites In: Proceedings of the 29th Symposium on Operating Systems Principles.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of the 29th Symposium on Operating Systems Principles

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.552954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:60fd09e1af52e188182b86d8664d89788afc286a24b444ff455e538aec3227b8

Observation 1b63beae-a6c2-419b-a386-ed53dc42d81e · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.448484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:4f671281a46232e33e3e4908b9c53350be1ce0ae4daa89d562a62ce6ecec6343

Observation b73ccd39-c036-4e3c-9561-e58abe5d4ad1 · outbound

This paper cites Let's Verify Step by Step.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Let's Verify Step by Step

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.451426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:cabe3a624d8bda7f2f9e51a7f2bbfe46187b48761ace6d088a81fa5cb914c8ed

Observation 612a62ee-1087-47e0-9942-d071da7629f7 · outbound

This paper cites Data Scaling Laws in Imitation Learning for Robotic Manipulation.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Data Scaling Laws in Imitation Learning for Robotic Manipulation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-29T02:14:00.368555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:61333c2f014cd790a5421c7bcad425c9bd609bb7fd5a39e90eab2d8768346729

Observation 41219db2-4d88-4017-b72a-a9c25cf6d625 · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 36 (2024) 2, 6.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 36 (2024) 2, 6

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.555326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:da0f658e5257b0eda7690744ee8f419a88fe60e0008efe08e3e4683588393371

Observation 6c2e92e3-1486-4928-bdfd-ef75be74d8c0 · outbound

This paper cites In: Proceedings of European Conference on Computer Vision (ECCV).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of European Conference on Computer Vision (ECCV)

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.557675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:144ad8904e82d025a9c7bbe5bda808ff4c9ebac8be092eb8af002a111f74981e

Observation d3fe87ec-04d6-4d2d-ac6c-958d11db9626 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Understanding R1-Zero-Like Training: A Critical Perspective

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.464329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:7daf73345b8bff04a752b4276511e80a29b9f46d59e84e23ea8f9ae4156fa6d0

Observation cce31cc0-399a-4cbf-b173-dab8a235e16a · outbound

This paper cites ThinkBot: Embodied Instruction Following with Thought Chain Reasoning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning ThinkBot: Embodied Instruction Following with Thought Chain Reasoning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.467721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:5ffd8dd4aaaae12db269a764587959872d87755acf6a4a3659beef0e53fedabc

Observation 68a59a5d-e582-464c-87bf-7959dd378f3b · outbound

This paper cites AW-Opt: Learning Robotic Skills with Imitation and Reinforcement at Scale.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning AW-Opt: Learning Robotic Skills with Imitation and Reinforcement at Scale

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.471034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:29ae8c4d612b9c6bcd8f391d3c2c9efc2e7c8ca39c83bbec75778787ac4ce3cb

Observation b2a778f1-2dbb-4a81-bc48-586505eba148 · outbound

This paper cites In: IEEE International Conference on Robotics and Automation (ICRA).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: IEEE International Conference on Robotics and Automation (ICRA)

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.560500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:3b136f7a2456f8912dc8618a19dbd3a57333ddb050f2b0911619f218b07e99e0

Observation 77728afb-05ac-4ec6-ac06-10df5faab142 · outbound

This paper cites Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.477240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:944e61b257b79d4cd848d47523615ee81ecaa163554820da304a60762de43d78

Observation c5c8edd8-14fb-461b-82fe-f6c95731e838 · outbound

This paper cites In: Conference on Robot Learning (CoRL).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Conference on Robot Learning (CoRL)

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.563327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:fe6024232acf93b04a7278e2046ba98e66e809ac2830f62aadea741ae36fc0c9

Observation fdeb00cc-2067-4a65-a583-ffdbde5915f9 · outbound

This paper cites In: 13th USENIX symposium on operating systems design and implementation (OSDI 18).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: 13th USENIX symposium on operating systems design and implementation (OSDI 18)

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.566046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:2e05de25ec880c7e9b7b3e34c0c94d0ccedb67d30b1b621ff2793e07df38c48c

Observation 9577ce28-66b3-4add-92d4-359e14dbd690 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.486956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:8a77dd96855f5421e8fd6d8c793e22ea4e7dc59d3fe3d1003240e85e4c37134b

Observation d50dc313-3588-4bc8-ba8c-80acc68977f3 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning DINOv2: Learning Robust Visual Features without Supervision

Reference 55

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.490754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:1369c073f622f8a6b0bae9ec34226ba78af7bf0f83431c6a39159011746990ad

Observation c8f2ce38-05c1-421a-98a7-203f9b9b970f · outbound

This paper cites Training language models to follow instructions with human feedback.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Training language models to follow instructions with human feedback

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.493812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:07b39f14362c7fcd59c1b7d9ccddddd542a73f79cb534697c3f6f7a54750759b

Observation 6990683a-0047-4f84-bd40-d34b0f706960 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 57

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.496669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:15e1247200c052b9a3df13944baaa18c3909015244e5157044d3a66351737c15

Observation c74546cd-dfe2-4a6e-8bf8-bc47efdae9bf · outbound

This paper cites In: IEEE International Conference on Robotics and Automation (ICRA).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: IEEE International Conference on Robotics and Automation (ICRA)

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.512174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:14e17c316eeacc92d42a3f0236993aa5bc7f3f3fe9f023dab044e2411c21a8b7

Observation d9b6d6c2-4ff2-4387-b4e8-ac82a01c600e · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) (2023) 7.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) (2023) 7

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.570595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:545f0c344c3e384021aadb2044e7eba0cf8bfa71e12bd1c05c2b31f26bef1e00

Observation 07556621-1d89-42b3-9560-e09a19f183c9 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.505914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ae7cc518bd1ea9d3e1fa5c8e10284b44c3c0647932f114debb8b01c86382cc5e

Observation 916f6c37-2c5a-4aed-835e-1b1828d8a4ed · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.509461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:f5e6d84225b6def3ce11ef2eff9d3038782408875fda7deccce99cb44a40caa7

Observation ab42983c-4248-4c9e-9466-bf77727e168f · outbound

This paper cites Proximal Policy Optimization Algorithms.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.286333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:dac31f821f4dbb74348741016d9a389a09dc14c60be3e95bb693a1082d2225ac

Observation a7a9c857-8ab5-48aa-bb77-f21c2f6cdb9e · outbound

This paper cites Google AI (2025) 1.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Google AI (2025) 1

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.572632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:b3c934a89617b05a1b60c7cbbc0d2e39cb857539737e092b2449b940b8f42b0d

Observation dcc32ed3-7850-4de0-8b33-3c4bf84624f3 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.297746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:cad5d70ec025b2b7d75ee368d1f1ebc7bd900f9de47ad2d13d11303194e31c46

Observation bc646701-724d-4f42-a316-0d034691272b · outbound

This paper cites In: Proceedings of International Conference on Learning Representations (ICLR) (2023) 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of International Conference on Learning Representations (ICLR) (2023) 2

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.574726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:7da0ff62c003fadcc0dbade2b5ed144c28b461c284f10eb5cbbf479768ec1236

Observation f71efbaf-08a7-48ce-99b5-42145d35d2c2 · outbound

This paper cites Deep Reinforcement Learning for Robotics: A Survey of Real-World Successes.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Deep Reinforcement Learning for Robotics: A Survey of Real-World Successes

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.305488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:628853ded13ca0ec9eea349b3702b8b72d75f2033079f792a38111b2d7c3d722

Observation bdaa6fca-9357-476e-809a-981017fc675f · outbound

This paper cites Journal of Machine Learning Research (JMLR) 10(7) (2009) 2.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Journal of Machine Learning Research (JMLR) 10(7) (2009) 2

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.577062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:d298feae3b61eef9af7a1d6fd8778482bd00da7d7186d9cbe5f11b61a814f0b0

Observation 54d36f7c-67f4-4d00-998d-bb395c545405 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Octo: An Open-Source Generalist Robot Policy

Reference 68

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.314752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:15a798178f204f0e7e146ba0c9a865cb11b60669089d97fe15985ff8d40ac157

Observation d2bbe571-76eb-4455-be43-3dc21c6908be · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 69

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.318456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:4d150e43ae10138611071554b107b59efdd4eed07a05d7e6cbede2bc2887e507

Observation 163e1871-2b92-4f73-b0fa-bee9785e5baa · outbound

This paper cites In: Proceedings of International Conference on Machine Learning (ICML).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of International Conference on Machine Learning (ICML)

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.579446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:8af0b8638566049a0f68e4b752a1530933f0f454327024b9609a07e4dd083026

Observation 8dc8bca2-8309-4809-81fa-feaf670d7e7e · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.327647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:920d615171b05ffe245c91910e4a04dec16f7e9f21be3cd0df23f2efb85fab5e

Observation 90a8e952-af5c-459f-996b-d7390064a253 · outbound

This paper cites an unresolved cited work.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-05-16T12:55:40.581579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:7128d6355c59fc2e3d81610ac956c678ada217768d141ffdd52fde799e97d75e

Observation b41fad6c-2b33-46c5-81aa-f41d4b4846b9 · outbound

This paper cites Hierarchical Memory for Long Video QA.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Hierarchical Memory for Long Video QA

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.335129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:f5a6c7d4bb0c31661983078e5dc54278d931135d5d805bb392a33ed0d05b9e49

Observation 23ab67ae-3c8f-4ce9-8b83-4fad3e7eb571 · outbound

This paper cites Ponder & Press: Advancing Visual GUI Agent towards General Computer Control.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Ponder & Press: Advancing Visual GUI Agent towards General Computer Control

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.338585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ab2a7e25c45b24f05281da863d0d861c9e4ed10a4fe32d184e639bc3ab1dc646

Observation 5ab7702e-0e98-4869-a4c7-6d183698e938 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 75

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.342411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:568d5ef258449ad503478a8ca734998701202e5a75110c903804f2475291ca02

Observation 57a908c9-f339-4412-9588-70200ab91d19 · outbound

This paper cites Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.348097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:5a1b253fe5b5fe49cac12015c9b83726ee4e3ac6f8d62d20b912aa0401759447

Observation dfc0f75d-3459-4f9e-9f3c-2dd025d7d7a5 · outbound

This paper cites Bootstrapping Reinforcement Learning with Imitation for Vision-Based Agile Flight.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Bootstrapping Reinforcement Learning with Imitation for Vision-Based Agile Flight

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.353532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:772e256d374931bc3852d916c67187bfe52685312b9ee999a715ffc4f289d513

Observation 3364acc7-9aa7-4326-8492-fcdad8c6e926 · outbound

This paper cites RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.358784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:fec5989461a611cf6174b269b116625e2512ab7b7f498b121311f1e4834ee038

Observation 4e11f906-e6ad-400f-ba18-078970dff357 · outbound

This paper cites Qwen2.5 Technical Report.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Qwen2.5 Technical Report

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.362546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:dfb6fa83e254a432c59ddd703bd0e39677e5e85639496529453f19e75699d875

Observation 5f9f9108-cba8-4174-8c38-58b4ea014579 · outbound

This paper cites ATP-LLaVA: Adaptive Token Pruning for Large Vision Language Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning ATP-LLaVA: Adaptive Token Pruning for Large Vision Language Models

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.366251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:9dcbe34dbe4d14b51576ead14cc83dfaaa7068d439798bdf5532c93f5cf38127

Observation 949cbecf-4b31-44f1-933a-4ae41c6f8d5f · outbound

This paper cites VoCo-LLaMA: Towards Vision Compression with Large Language Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning VoCo-LLaMA: Towards Vision Compression with Large Language Models

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.369819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ea13437e503ec3017cae3173e69d358b4d68542327153f30bf6a4395bc155bc9

Observation ad890374-dd81-4fc5-9731-dd33c7d49399 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-05-16T12:55:40.372850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:158172339df24dd60fee1650b801e57fbe7ea7f594720ee04e10e488ae917645

Observation 9a5971cf-3333-4241-a5ac-91d801969469 · outbound

This paper cites Scaling Relationship on Learning Mathematical Reasoning with Large Language Models.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

Reference 83

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.376599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:c2fd904d8673b44c2033697c5dec5458f15923a7f18719015f6fa964105b5458

Observation f808a29e-9e89-498a-903b-3865ee98f85a · outbound

This paper cites Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 15476– 15488 (2022) 3.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Proceedings of Advances in Neural Information Processing Systems (NeurIPS) 35, 15476– 15488 (2022) 3

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.583767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:10915a142201f990bcbb520ccf7f765df73d85cf794f37f7217e91e4a7ebfa58

Observation 4db77e89-7dfd-4933-9eca-fcb3fc8ac11c · outbound

This paper cites PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.383944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:ad13435ec853197a540f8928350ef5bb54e3427298b011bc33ec5c4da2cf8bd9

Observation 43afe699-a7ef-47e6-8ac8-4502143488b1 · outbound

This paper cites In: Proceedings of International Conference on Computer Vision (ICCV).

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Proceedings of International Conference on Computer Vision (ICCV)

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.586037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:a1eeb407f997912c7ab6f488f3b2eaa3c58257046ac3345f39ce239c19c5367e

Observation f614c2f9-c1ee-4bd9-92c9-241b1d2d38bb · outbound

This paper cites Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.390683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:df568681aff4a951b55a0e5ca6fa820ba320205c8d644f56c3e1ec06176c9793

Observation 586d7973-c27e-45c9-82dd-b467357d334c · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:43:43.298530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:345e5e1920ed419002dc9033e87a96d8250c73b02f619f99b54d6e0e46fd00ff

Observation a50457bd-f4cd-4f9a-a775-a7e4dfb9c6c6 · outbound

This paper cites GRAPE: Generalizing Robot Policy via Preference Alignment.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.398126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:7f9266356cd846cf3403e309b7c30543ca6d392fbcd7f626e8d949dc0fd5e511

Observation 65a0e00b-9d5f-44d8-9679-92277947ce43 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 90

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.402013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:5ee163d04ceb5d7ad9fe4328e8ed3a21eda142adfc2f5ce66d3c326ef2025f02

Observation 7b83022a-a003-4080-a99f-2f33b4a44360 · outbound

This paper cites The Ingredients of Real-World Robotic Reinforcement Learning.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning The Ingredients of Real-World Robotic Reinforcement Learning

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.405729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:a04d085b0df50b26890f13cc1f81fbeb58d3e99aa306ace608486d56969e60d2

Observation 001293e4-1d88-448e-ae58-dfaf748352b1 · outbound

This paper cites In: Robotics: Science and Systems (RSS) (2018) 2, 3.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning In: Robotics: Science and Systems (RSS) (2018) 2, 3

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.588575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:3484387f53d9547b0e944367eb16e601877a14b351a0664ab5fbca41bb157e93

Observation 0632fdfa-b1ed-42bf-9b25-24c0441a7d65 · outbound

This paper cites Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2023) 2 15.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2023) 2 15

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:55:40.590806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:1b105f7e281a5fd5ea7aa966e2c661be4a308f99487c31d39084abdc0592db50

Pith citing papers

Observation eca419b8-08fd-41a2-847b-86be3a85d6b5 · inbound

RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models cites this paper.

RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:36:13.013125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:36:13.013125Z digest=sha256:4a4189f74845b0d9573b929c08061717834caeb78000cb06c3ae221e612d19af

Observation 95ea6872-0446-4211-8b8c-e7c1baa73fb1 · inbound

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey cites this paper.

Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 181

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:28:16.127771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T20:28:15.818016Z digest=sha256:522549b354169777a9cac1028b007be45c044f21ccb4562b2925d975bfc56b35

Observation 7cc1bd6a-b221-4f54-b7e7-2c0543864e5a · inbound

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models cites this paper.

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T10:30:23.605483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:30:23.605483Z digest=sha256:e9c2f7aab10bbf48554a4f230219bc2b4539fb17dc9fc8b14ebb9e05a61e80f2

Observation bfb3efbf-d878-4a51-ad43-ad1dee69dac5 · inbound

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions cites this paper.

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T12:42:46.978148Z digest=sha256:578b504d33e11efadead45c1ade8c3e3032164bb985ab7fe622ac95a85584679

Observation 97d017ce-8497-4e0b-8cd1-d2c3c6eff92c · inbound

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning cites this paper.

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T08:02:11.189795Z digest=sha256:498ab8f850cb7f5a257aaca27a0b42d3225da24f0fd437fe41cf3de332e90177

Observation f4cc08e2-a87f-432b-ab5f-013e68fd6365 · inbound

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning cites this paper.

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-21T21:34:22.289168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T21:33:40.229376Z digest=sha256:53e00eb9ca97c05f60da3c414b914f70b523aac74ce56fc74088fb89cc56172c

Observation 1ac2bdf5-083d-4463-beb1-d9c17d00f5f2 · inbound

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization cites this paper.

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T06:20:01.988199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T06:20:01.885711Z digest=sha256:4310343c4ae4c0398db14b9fd48f7e21fe3cb0195e6d9f396d3a5bcaa0a57034

Observation 985939ba-55d8-42ff-9a0a-150adb4c5433 · inbound

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization cites this paper.

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:38:54.514977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:38:54.514977Z digest=sha256:8b27fefe8bcc79311782ec3b3205d903ad72c9c03fa9487e3733d868bb2556f0

Observation 3a7f4c9f-d4c4-4287-b44e-bfbaeb708d9f · inbound

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models cites this paper.

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T10:24:55.949494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:24:55.949494Z digest=sha256:87ab383982a734868a79ffdf64c44879a1f5c864f43182e32b044af72b51c314

Observation c3dc4031-dc1f-441d-906d-44316b796107 · inbound

Reflection-Based Task Adaptation for Self-Improving VLA cites this paper.

Reflection-Based Task Adaptation for Self-Improving VLA VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-18T07:31:02.950270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T07:28:11.187479Z digest=sha256:690ccb698e2eaa9bd1819c68531ea078363a2cc0985f8743f180bcea623cda81

Observation 45c86c92-7ae0-4ca6-8b7d-af521de763d1 · inbound

RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation cites this paper.

RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-18T06:10:57.839817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T06:10:47.309028Z digest=sha256:2f5795290ad0983b8866d7dde51a7d14b05ba087831b8a99d9305456726eee6f

Observation 18c063c8-8a2b-4822-85b1-e981fe93cbcf · inbound

$\pi^{*}_{0.6}$: a VLA That Learns From Experience cites this paper.

$\pi^{*}_{0.6}$: a VLA That Learns From Experience VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T10:34:59.134604Z digest=sha256:b907a56ec48f9e489d60318b05372e8b7bb109b30a7d913a78acb87c66f5beca

Observation 39b4ac06-45b2-4ad4-9346-891b67d65b1a · inbound

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models cites this paper.

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-18T03:10:48.847422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-18T03:09:09.713822Z digest=sha256:d76abed27a18423971afae82b2e6ae6346f3079d1df464dc226a920edb80014c

Observation faf27b8d-eb4f-4e81-af34-248ece198293 · inbound

ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models cites this paper.

ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-17T06:09:09.446324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T06:07:42.311608Z digest=sha256:c86616098d8a8603d901f00eeaf60435fd7db2d1420d18359c5ce1b5afe74443

Observation 530a6ef8-e99d-4078-bf48-544523076412 · inbound

RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training cites this paper.

RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T07:01:54.696390Z digest=sha256:00fe521f6c9186fc85ed06edf9785e0059d4322b81c13589a97448bb90f120f9

Observation 4f9935aa-8b8d-40cc-9b71-483af13c20c6 · inbound

World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy cites this paper.

World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T03:56:47.172497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:56:47.172497Z digest=sha256:7a0a9bdbbb0c8a289c148650d6baa2a2efbe567186dc3ba1059b89ade082155e

Observation 6a0d8fac-a5cc-409a-a21b-43b808db6e14 · inbound

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation cites this paper.

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:14:10.829151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T13:13:53.818915Z digest=sha256:a5c3a40fe399cb777b25621291b4edbe45b7ffd2dc580f3ed91754ca3f2205d0

Observation ea9a4a92-774d-4402-8c7a-0038384e2946 · inbound

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning cites this paper.

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:10:13.210273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T14:07:10.387869Z digest=sha256:0bbdede3d6673b4dad4b0c1da9670bb3f699569f8e1010e3d08819cc6146c742

Observation fc388521-7d39-4880-ada3-0e86fdca0b00 · inbound

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models cites this paper.

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T06:01:30.803128Z digest=sha256:e519949ed9927b80eba8c8e2ea605935cf05b6c3cad62d91362a45308fe42164

Observation e2d9f503-7dbc-4b8a-bfe0-f781d72545c3 · inbound

RISE: Self-Improving Robot Policy with Compositional World Model cites this paper.

RISE: Self-Improving Robot Policy with Compositional World Model VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T02:28:37.997148Z digest=sha256:d220cd66cfff8398abbee5f3cff10376d3b9d08570e083f727fd3deeceb2ae58

Observation 060211d4-f4da-40b0-ab13-77d5039ca6a8 · inbound

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training cites this paper.

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T23:47:56.865774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:47:56.865774Z digest=sha256:26dd5f1ad866c423df19469feba20f2b330c02553f75f51355577989b2df7c1e

Observation f7dd5d3a-bad9-4576-b747-d22c80cb96a8 · inbound

WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL cites this paper.

WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T23:23:52.056025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:23:52.056025Z digest=sha256:00c91f874bcf8060cf0c9f902402dcada07ccb1e5c38b6c588d331210eebf0a6

Observation bb374873-0a04-4ebc-b858-0cd97984f424 · inbound

VLANeXt: Recipes for Building Strong VLA Models cites this paper.

VLANeXt: Recipes for Building Strong VLA Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:00:09.810733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T12:58:30.777235Z digest=sha256:cfc7d7c6aff7ed65143c8b3de5a7cd84f5d9d0f526add7c438805da6de15ac11

Observation 71402313-8f76-4efc-a5be-c6222c35c19a · inbound

Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models cites this paper.

Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:48:52.392191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:48:52.392191Z digest=sha256:ae6f262d3a341475cc9905b92475cc7b3dbce7c36bfdef65933ec240deefd267

Observation c6932cf6-92e9-4687-a09c-c077f07da32c · inbound

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector cites this paper.

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T09:53:45.151347Z digest=sha256:97011e9c2fd8a5c93aaf350a7c930473d1a0c5e2f34d8aa1e2651b7e4d57ad14

Observation 53f5ac62-5448-487b-900b-019c9836294a · inbound

Genuine pair density wave order on the kagome lattice cites this paper.

Genuine pair density wave order on the kagome lattice VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-13T13:17:06.175932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:17:06.175932Z digest=sha256:9087c0477a6a0ac0cd6f2ce9666757580e929cc5c9a68484a69778c4ac70d398

Observation 1e6b6c48-e063-4011-a61f-480bd6c2a56b · inbound

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes cites this paper.

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T20:21:45.365156Z digest=sha256:f445aaa5aab24ed0959ff91bcbe452ea8286495ed2cc32eb1f13b3bec6ca0a6c

Observation df059891-b1b1-4286-840f-2eed7ad45c5e · inbound

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning cites this paper.

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T19:16:46.753641Z digest=sha256:571571eb697fcaac799cf553bff35867bec6a8df567303bf0f89e80f6fcc8039

Observation 09a50abc-79d3-432d-8864-bfc7c9ac95e0 · inbound

MoRI: Mixture of RL and IL Experts for Long-Horizon Manipulation Tasks cites this paper.

MoRI: Mixture of RL and IL Experts for Long-Horizon Manipulation Tasks VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T16:01:11.260956Z digest=sha256:c945373c39c7d077445d76d669ad0793739cb6fb1b74b5c113d75ece324c0d1f

Observation be4f14e8-f806-42b4-b890-95755c6f7577 · inbound

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL cites this paper.

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-10T05:16:57.241755Z digest=sha256:19c6f7d336b4ec20793cc151d6967e6f82a0fda64f4b9e5e143f171c1b77179c

Observation 16caa9cb-90b4-4183-9026-3738f13700f0 · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T22:07:24.208555Z digest=sha256:d3060b891a1294e40b9603d2ad26487e82d5394c6cb314a21e4a35075a8aff6b

Observation 938c5fd9-e0a6-460a-892c-4fe583d2a80a · inbound

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors cites this paper.

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T18:39:49.173828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:39:49.173828Z digest=sha256:0cb80b7a6776b478a0bf7c76391027be9f4e35aa4d50a70666777ff6668ba049

Observation ad842d11-fe5b-4896-95dd-4c9f93f4911e · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-07T05:47:17.494531Z digest=sha256:64b57dc5cc1bd7302baf4b6580c9d129430d41630df21bfff2e87622be264972

Observation ddf9aa26-5cc4-4740-a75c-80c8a46728f6 · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T03:00:26.352130Z digest=sha256:46fe8c4c973c286aa25a430349a28005a0c910f65c5a5bc2f5388b8f6e0a63e1

Observation 5d43047d-35d9-4f0b-b2d4-c2bfc2fdb059 · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T19:31:38.069592Z digest=sha256:1fda98b190660e087e66a0637f9cdc658e5939e411add47e4794081b6be794b7

Observation f4b8ac92-607d-4c3c-97a1-5471416adf4d · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-01T08:15:32.305693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T08:05:47.128354Z digest=sha256:f42d8a06ba152ee2efe70dfe28fbf5d686f81dd6991befc5aeed3b3364754b77

Observation 97e60cf9-582a-4d3c-add6-fca7f9bf2ddc · inbound

Closing the Loop: Unified 3D Scene Generation and Immersive Interaction via LLM-RL Coupling cites this paper.

Closing the Loop: Unified 3D Scene Generation and Immersive Interaction via LLM-RL Coupling VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T14:53:46.528750Z digest=sha256:c68984818a217fc29086da3a5299bc7a71dd134e2867ceda236cff78a6e24088

Observation b023fe87-8298-4c0a-b15a-bfd05d702541 · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T04:08:43.222818Z digest=sha256:c8ee839577552218aa4c9017c394d42959e17645ffa139b48db1c0c360ab1bdb

Observation f980e6c3-9153-44e2-a584-a2759dc964df · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:01.263542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:01.263542Z digest=sha256:cf1a7486e4bbe44ade482dbf0ac21de0d4c8f2ee0b8e3341c280a600d3636dc9

Observation 51e9b58d-dd57-4457-a838-e13fe2cadfbe · inbound

Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation cites this paper.

Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T05:25:18.120832Z digest=sha256:7f7c52de961cebd17d1ee3bcf377eeca4cba31d4e2142c85680d5f0c9bd014b2

Observation ccf925a0-4a78-4ea3-9ece-2b27b9772ded · inbound

Reinforcing VLAs in Task-Agnostic World Models cites this paper.

Reinforcing VLAs in Task-Agnostic World Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.591564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T04:17:51.349213Z digest=sha256:1df71a48d745b32b3d7ae8c7104d213af0ca407518162935ca502853c0b3b4c7

Observation 3eadb395-ff74-472b-86e7-b2924a69d19c · inbound

Reinforcing VLAs in Task-Agnostic World Models cites this paper.

Reinforcing VLAs in Task-Agnostic World Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-21T08:14:03.147881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T08:13:40.975340Z digest=sha256:5f4f9c357623b00041a265419ba0d43a94b580e1f3cdbef0de8a2fb08067ba96

Observation 1455a558-33c1-41c5-84a7-6f441d26eda7 · inbound

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking cites this paper.

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-20T21:09:02.704069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T21:05:45.024226Z digest=sha256:cef1006834b572101f81498bc7d1bd45e18293dab5509705c42200f96968569a

Observation bc55fd09-e136-4a2b-852c-2c7435e9a8a9 · inbound

DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization cites this paper.

DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 129

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T12:43:16.978981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-20T12:39:50.004269Z digest=sha256:1da1d8d2449bb54d74004f28b00ac556e249eef55c01d01cca34afd2052d20cb

Observation d7845171-13e2-4158-8031-77b30695902b · inbound

AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment cites this paper.

AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-20T12:53:17.670684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T12:49:02.418663Z digest=sha256:8ef7146071697a8e1cf04be9513201405a05f31582ea569237411d435f88e196

Observation 56dc571f-f0fb-4aed-8b64-3999e63c1f19 · inbound

RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models cites this paper.

RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-20T04:53:05.074152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T04:48:39.069675Z digest=sha256:0c281cc6760dbd6e96155cc8f153e673faad33cb0ff4008ef51543605987e74b

Observation 91e39c76-f4ea-45cf-a5ee-df5d0d8dc74b · inbound

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models cites this paper.

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:50:23.587617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-25T05:49:15.137844Z digest=sha256:519225ef0d2047ca379be926ffa291e6dec8ce6e1736b82d70edb89f2350e29c

Observation 1bf0a3ca-a72f-461b-b323-7cab20d2cdbf · inbound

EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models cites this paper.

EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:13:59.923584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:10:08.682307Z digest=sha256:b26080498dea0a6bcc0357edeba685a826512fa7591ce3ff18daa01197839f3e

Observation df1dc87e-2b68-453a-aad5-bb4965f5b013 · inbound

Ratio-Variance Regularized Policy Optimization cites this paper.

Ratio-Variance Regularized Policy Optimization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-29T19:53:55.988110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T19:44:02.317303Z digest=sha256:79d67587321f6317128c33ea0eb057c1dae880f3c65fc9ffc900558491b581fc

Observation a706de2b-40e4-4b2d-851f-ec8ccd2fd7f5 · inbound

GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation cites this paper.

GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:23:44.482736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T17:22:40.258560Z digest=sha256:d0a70f496aaa6d508d85670bcffd5b2e753cae4b2abacf284dbd26606bfd5073

Observation 92fe4348-2f02-45b6-bdcf-6a76e2e55036 · inbound

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO cites this paper.

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-01T23:16:24.703428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T14:28:23.759920Z digest=sha256:571a33b31c0cf89187dcded9637829c6997cdc9b4d4767e5ac16ea846ed0ebdc

Observation 66141af6-a326-4bb8-94bc-a52f2edd3006 · inbound

FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization cites this paper.

FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-02T09:06:49.369951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T05:38:11.089753Z digest=sha256:a9ab7377792c34e3a7dabdb93e876e810759dd1bf73a8c0b63e8ec3b170e34eb

Observation ecdabe67-1554-4e5e-a184-d74d9e48a83c · inbound

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning cites this paper.

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-03T05:07:39.010306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T13:25:59.194721Z digest=sha256:a7befe05ae532e378b65c1cf461c0edba36bf9c9df48f92a5b6221e0521c3c9f

Observation 3cdf9c80-2550-4181-8045-1642bba7be43 · inbound

InDex: Empowering VLA Models with Intent-Conditioned Arm-Hand Coordination for Dexterous Manipulation cites this paper.

InDex: Empowering VLA Models with Intent-Conditioned Arm-Hand Coordination for Dexterous Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T11:38:05.561555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T09:24:25.659815Z digest=sha256:69f7962d5c1f3c5620da991eece44451f1f49fc5bc33ad3871d81f1a3692fbd3

Observation 18b9c67c-fb9d-4a1a-b5d9-c3fe6e0e6c8d · inbound

InDex: Empowering VLA Models with Intent-Conditioned Arm-Hand Coordination for Dexterous Manipulation cites this paper.

InDex: Empowering VLA Models with Intent-Conditioned Arm-Hand Coordination for Dexterous Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T11:49:12.610001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:49:12.610001Z digest=sha256:1ad1b41b284d64a39af578ab40e8e2c3ce404c5dab80535b8eeedadd1763d928

Observation 2d07dac2-23cd-4b0f-91d8-0c274b787cab · inbound

PolicyTrim: Boosting Intrinsic Policy Efficiency of Vision-Language-Action Models cites this paper.

PolicyTrim: Boosting Intrinsic Policy Efficiency of Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-04T09:09:43.337115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T10:27:00.283283Z digest=sha256:39469d18b325ce90cd25ca45c78e33aae25a9271ec8ea90ca94ae233963f76b8

Observation 80944bba-9465-4222-b737-d8dd66f6f9b4 · inbound

dVLA-RL: Reinforcement Learning over Denoising Trajectories for Discrete Diffusion Vision-Language-Action Models cites this paper.

dVLA-RL: Reinforcement Learning over Denoising Trajectories for Discrete Diffusion Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T10:59:46.958851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T08:11:41.429127Z digest=sha256:599a62b6169cd5f3994896d7bd3a53bee192a081aef30babed920fe32182f74e

Observation be122517-b9fc-41a9-9dc5-df9c8dc11f55 · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-04T09:59:44.633156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:ae447da87eeb7c38f0e0dbb2222852ee33bed900ae99fa5efb9191772757d11c

Observation bcb2f0f8-2372-46da-896c-f3e0b99b4b0a · inbound

FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation cites this paper.

FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T20:50:12.338112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-25T19:29:00.117285Z digest=sha256:115415934ca15e9972cea68b25aaf72d3a603de0cb88e141ca90cd6e973658a0

Observation f10a5cd0-a4f3-4e14-9f7c-ce946f10cd2a · inbound

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models cites this paper.

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T06:04:21.080240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T06:02:15.781538Z digest=sha256:f82508c1498d2ba797b4c0c63c0ce701368d9015211651c6e69f2f7bcc668117

Observation f4a4d3aa-fc04-48bf-968a-5a2ff7b39904 · inbound

Adapting Generalist Robot Policies with Semantic Reinforcement Learning cites this paper.

Adapting Generalist Robot Policies with Semantic Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:45:42.701388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T05:09:29.625066Z digest=sha256:ac495995431ded9e49467e81fd0353ade70354f24a3ee2f685553eefd755d97f

Observation c01e0230-32b9-4630-abe2-a37dba55aa79 · inbound

WorldSample: Closed-loop Real-robot RL with World Modelling cites this paper.

WorldSample: Closed-loop Real-robot RL with World Modelling VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:58:02.484283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-03T10:57:40.128651Z digest=sha256:fab45f8443eca27ee993a8e593d87e24401d2d196b418fc79a1bc71127326dbd

Observation 4fab7b81-c372-41c5-bab5-ab5298d2d764 · inbound

HALO-WA: Hybrid-Attention Latent-Guided Online Reinforcement Learning for World-Action Models cites this paper.

HALO-WA: Hybrid-Attention Latent-Guided Online Reinforcement Learning for World-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T20:32:22.412216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:32:22.412216Z digest=sha256:f46b3279989b870380fa9c1d31eb1f105d17120b2d7aa706c9f5ad146bca82c1

Observation 09598840-1f7d-4b8a-afd5-42772da2949b · inbound

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review cites this paper.

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-07-10T23:07:47.946348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-10T23:01:01.563768Z digest=sha256:57976131255a1bd5a4c2fe7557b28df907b90ff7de2ed4ced88dca720db7be57

Observation 1c410481-3d33-49ef-979f-45376a425704 · inbound

Prompt-Driven Exploration cites this paper.

Prompt-Driven Exploration VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T06:21:42.026251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T06:21:42.026251Z digest=sha256:bed268768740f092d0593444b8e8566d6fe302b1e244f5c52774a1284f7b28d8

Observation 379bb78a-3c73-4c7c-a487-415db58728cb · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:39.365484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:39.365484Z digest=sha256:9574f9c34dbed7b170cda5621302035c481169e6e84a9ab758c6cce0f4177368

Observation 1a3c64a7-deb3-48fc-93ae-938cf8fcc637 · inbound

$N_0$-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens cites this paper.

$N_0$-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-30T12:43:44.568407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T12:43:44.568407Z digest=sha256:53e60ae7581029e3bf9f79470e1f5d8d48bef4f4f29b5fca0662514ac98a4588

Observation 01bf5c1e-2ec4-4feb-ab87-83801cc30ddb · inbound

RedFlow: Redirect Failure into Action-Level Corrections for Flow-matching VLA Policy cites this paper.

RedFlow: Redirect Failure into Action-Level Corrections for Flow-matching VLA Policy VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:24.481776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:24.481776Z digest=sha256:dd29eefb393f7e575f8ea3cff05112d66103f2c8e33b9d968ec6bb44d26be5e7

Observation 123f041f-05e5-42a0-9235-954ec290d52b · inbound

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning cites this paper.

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T03:34:03.842090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:34:03.842090Z digest=sha256:11f98c425551da223b956127504d62690dc9102b2cccd7d24d793ef81d816972

Observation 884978f0-2f9c-4d2f-8b90-121652174f19 · inbound

Transforming Remanufacturing Automation with Large Language Models: A Forward-Looking Analysis with Case Studies cites this paper.

Transforming Remanufacturing Automation with Large Language Models: A Forward-Looking Analysis with Case Studies VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 176

Resolution
unresolved
no resolver link, observed 2026-08-06T14:54:47.953372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:54:47.953372Z digest=sha256:8a165c9875f3e35fc531185ec00d9df89937aa19afcc38ae01f66b9289338e83

Observation 9ce8dffa-5f44-4db6-91e5-24620bba3f4c · inbound

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation cites this paper.

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T19:38:41.738287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:38:41.738287Z digest=sha256:0b6d172ef18cecc1f2375e4df204291101b0b28ed6128c9f5963a257d2c815f9

Observation 5476859f-65e1-4d2b-adcf-a043a72076f6 · inbound

TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models cites this paper.

TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T10:29:55.257629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:29:55.257629Z digest=sha256:aecd3e789bca430d8a05329f31e7c8d2d04d61a4d74cc2569c10b5e2606918d7

Observation e1a778d8-91c5-41fb-b680-01ea895e6215 · inbound

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding cites this paper.

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T00:35:34.159303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:35:34.159303Z digest=sha256:6ca7a8f30826682f4b08c965e6ff1516ec77d44c7f224bfe0e75b4f48e5b643d

Observation 8fe47e55-fee9-49fc-b126-586f27ee244d · inbound

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition cites this paper.

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T11:20:19.953071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:20:19.953071Z digest=sha256:c34e067e0d7f559aa44c4ea4ecdd934e7fd0116543eba49b39c57f8c9a91c483