Pith. sign in

Paper Citation Record · LEDGER

Improving Vision-Language-Action Model with Online Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2501.16664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.16664 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:40:21.447558Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.548397Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a4dcedf2-a314-48ff-a734-3fb524645888 · inbound

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations cites this paper.

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:38:11.188770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T18:38:11.110166Z digest=sha256:e8caa001f860c8a747d5be34bca26dc6735141c1bf229d3a11cd05e528b0610c

Observation b0f97b08-27e5-4fcf-a420-409b0c1ef202 · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.815167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:3b395a85e588ede36bdbf119441c9058279ecf0d50a119eecacde9b95e5b5747

Observation 7108ed8c-d69a-4c54-b1fd-20506952726b · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.499761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:82d0ed831fbb364bf161b2b3dc70eaf3e26f238a7959b0036bf078381c23cc6d

Observation 7c594d43-1aa7-4147-8322-d05287c928d0 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:21.447558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:21.447558Z digest=sha256:452e1aaa3bc66cb56174df4e3265b62151ad8fa62c7dcad591dc70da116c6f1d

Observation 65301ea4-71ca-451a-a691-056ceb3600ae · inbound

RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models cites this paper.

RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:36:12.864759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:36:12.864759Z digest=sha256:6bf9e649958e0f14b290894deec9e3e37b8cc1363f3c4896718e672007fca90f

Observation 689c3033-cf16-4782-a3a7-eb31f2b0ae21 · inbound

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training cites this paper.

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:06.290688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:06.290688Z digest=sha256:423d04f0b4d126120a4b326ead31a8589db04a0e3b1b89de0cc398176abbbe00

Observation 2634ff7d-e50f-4d87-b330-b96d3689fa0c · inbound

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach cites this paper.

Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:35:43.638797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:35:43.638797Z digest=sha256:a8c4d48ca2da152947b76cfe4452ebc39029b766c197205b49d835e9b0db81d2

Observation cd37d08d-cc3e-430e-9226-c59979728f02 · inbound

Leveraging OS-Level Primitives for Robotic Action Management cites this paper.

Leveraging OS-Level Primitives for Robotic Action Management Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:38:40.022379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:38:40.022379Z digest=sha256:9581ee2a7eba79d9309322767bb193482101b8795c41da1b971d876e803c94ec

Observation 35ba3196-ca77-4c1e-8e83-1549cb75e2ac · inbound

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models cites this paper.

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:30:23.538701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:30:23.538701Z digest=sha256:344d679afb9aa320f7ee601b9c87468a858bae2bb9cfdcf59a01270ddb8a9c7e

Observation 2725aca8-5c7c-48f3-8fbf-57cd903c3e8f · inbound

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization cites this paper.

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T23:11:31.875241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:11:31.875241Z digest=sha256:0424ce5f5310047865e603949a8b992f19bd03f9b741e52783edb5941812db7e

Observation 0bfe0573-d8ee-4686-9445-2942dda656c4 · inbound

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning cites this paper.

SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:02:11.482284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T08:02:11.189795Z digest=sha256:13efa28bb683b797df16ff5e6872bd5f668e88ec892f5b321ce31382f8a6697f

Observation 1ab1b8f3-bc2c-4628-bbb2-d0007fb19793 · inbound

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models cites this paper.

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T10:24:56.027069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:24:56.027069Z digest=sha256:228cf14d418ac0bbe8f51b7c7b1e6b3e5c5e17aea52918c405bf22462f7f4fbf

Observation 98b2c27e-2eaa-4d5d-bc6d-f8c540b37840 · inbound

Ctrl-World: A Controllable Generative World Model for Robot Manipulation cites this paper.

Ctrl-World: A Controllable Generative World Model for Robot Manipulation Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:14:10.368029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T01:14:10.174044Z digest=sha256:fcf41655a45ce56fcf35e6479ed19a657bbda44aaf540f0f94fde52db3ed22db

Observation 90b52234-e401-4d30-be29-95df048fed96 · inbound

Reflection-Based Task Adaptation for Self-Improving VLA cites this paper.

Reflection-Based Task Adaptation for Self-Improving VLA Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:31:02.913540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T07:28:11.187479Z digest=sha256:c832c126408df78dc1f5a09f978af2df48a8ad5ef19af89afc76740590c6ed62

Observation 96f1fb71-e18f-481e-972d-a1fafab88cf6 · inbound

$\pi^{*}_{0.6}$: a VLA That Learns From Experience cites this paper.

$\pi^{*}_{0.6}$: a VLA That Learns From Experience Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:34:59.347237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T10:34:59.134604Z digest=sha256:eaa5fc3c47c3e779b1427dfcab3eb48e5aa074c065c872650dc2668383c15613

Observation 1ab87388-22a1-49df-b0e9-e41fa555e3f4 · inbound

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models cites this paper.

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T13:53:24.794126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:53:24.794126Z digest=sha256:5a8f1f7decbf224e7e79aad3096f068e97667881f0b44336acd7c007e2e295ff

Observation 50f1e294-9b47-479e-a108-da46695ada15 · inbound

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models cites this paper.

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T12:30:33.790260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:30:33.790260Z digest=sha256:faf1d4392eea97ffa13828f4b70f687a274314a6375ac927ebefb793c380696a

Observation 58ea4866-2491-401c-8be1-2517f1eb356e · inbound

VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation cites this paper.

VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:00:25.994640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T07:00:01.741166Z digest=sha256:ac4d549a6693fa8b8f80b098574303d2af78061fa5e8c4858b808540998906e3

Observation a8596dfc-c014-4905-aa7b-adc13a2adb55 · inbound

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation cites this paper.

TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:14:10.897478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T13:13:53.818915Z digest=sha256:474e21f6d45548aa20522aaea7e72bdddd63bb65073ed2c5a68a225475883c28

Observation 0d17fbb1-bf9e-414c-b44b-a41290644b48 · inbound

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning cites this paper.

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:10:13.288701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T14:07:10.387869Z digest=sha256:4d33689a88fc4f27ba2ec8e12f18599e76388c6d68f7589d84531ea62059b006

Observation 694260c5-fe28-4f0e-baf9-87fcd39c434c · inbound

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal cites this paper.

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:23:26.867422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T01:22:42.691009Z digest=sha256:4f9857e13b96c4520fe74b0044b568ee4e99e8ff298d783870b66aa0e51a4a3e

Observation bbc803a6-8cfe-467b-9d40-981a1f8da0fa · inbound

Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation? cites this paper.

Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation? Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:10:49.569291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T20:10:54.362107Z digest=sha256:9de7707df3827bcf00ff665f990d8c3d451ebc67d4fce061381787ffcbc89213

Observation f1ace14f-8eea-428c-b697-830e5a049e2a · inbound

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes cites this paper.

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:48.347579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T20:21:45.365156Z digest=sha256:ad567295e535bbae0d01e53958b3f886dc71d8066f099a36088d13c9c246c95c

Observation 44d8d85a-3225-42b0-bb30-1eb410e04121 · inbound

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking cites this paper.

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:09:02.659480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T21:05:45.024226Z digest=sha256:a996500734c6f51277c758bc28cbcc39741eb75d5edea7ecc0f15d4178375394

Observation 56c18ea0-13ad-4965-98ff-870fb0e1512d · inbound

DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization cites this paper.

DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:43:17.257279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T12:39:50.004269Z digest=sha256:f09f9631d6b3a660b0f4e9ce97fe1760d5e22c227b6cf82386a917adc34f6d60

Observation 510528f5-e4b1-4196-8011-37b2e07639e6 · inbound

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models cites this paper.

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:50:23.617278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T05:49:15.137844Z digest=sha256:e6f276a71c66566cd0b9f356828edbb0cc3dd38825422e8e56e5776b3da898da

Observation bbde08bf-32e7-433b-9ce7-afe9dc3ef770 · inbound

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning cites this paper.

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:39.001497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T13:25:59.194721Z digest=sha256:ee350747f398164c954f0edf5dc24c9dee7737d36466f3b59ed2409f7cbb08d8

Observation 69d21ad9-a07a-4301-af0e-5d08773fa56a · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:44.550081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:437cdfdcd346d0e87769c43ea08f051f1f00e416561cc3ad477997684e0bdff2

Observation 5a0155eb-bdf8-4494-bf6d-5b47aabf7ef9 · inbound

Adapting Generalist Robot Policies with Semantic Reinforcement Learning cites this paper.

Adapting Generalist Robot Policies with Semantic Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:45:42.654679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:09:29.625066Z digest=sha256:fd344caf5434e402256f6306f7b2c05eb89c30f04b463732451885a877c326d4

Observation 844fe8d3-167e-4ef6-a8df-f31242600c73 · inbound

RL Bootstrapping of OpenVLA-OFT for a Novel Robot Embodiment cites this paper.

RL Bootstrapping of OpenVLA-OFT for a Novel Robot Embodiment Improving Vision-Language-Action Model with Online Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:37:40.214321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:37:40.214321Z digest=sha256:9194a68770b85dbcc4030a6d4a50950cbe66b4155eedffb7cb4706fcad5d8030