Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:39:56.248989Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 34 inbound Pith citation observations for arXiv:2501.16664.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:39:56.248989Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:43.017171Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:59:44.548397Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4ca4a5c0-e137-4ee5-a4d9-5611a1d9aa9f · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Training language models to follow instructions with human feedback,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 34225ca0-7377-480e-964a-f12ea759e71f · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Improving alignment of dialogue agents via targeted human judgements
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da756296-10e2-402d-b8cd-d79a7c141e67 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning LaMDA: Language Models for Dialog Applications
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a3ca115-cde0-4104-a4c9-13334df3f68e · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b41036e-dd35-4d00-8dae-5a5979e4d840 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d02f40da-e6ae-4a6e-943c-5c771f8af09e · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Inner Monologue: Embodied Reasoning through Planning with Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 888e2de0-db23-492c-b225-5cc9bc0f05fa · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning RT-1: Robotics Transformer for Real-World Control at Scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 632c73c7-66a5-4494-abeb-292de82cfc4f · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edf33287-b221-4324-ac5b-f616fd3ad41b · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning A Survey on Vision-Language-Action Models for Embodied AI
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe8b26d-df61-47b0-b545-368be1f23d13 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Hirt: Enhancing robotic control with hierarchical robot transformers,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ac41728e-41a9-47c1-bddc-354ccd085d3e · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Vision-Language Foundation Models as Effective Robot Imitators
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f743d18-c3bf-4e2b-8ad5-8eb37e1df240 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Chain-of-thought prompting elicits reasoning in large language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 79c7c249-caf3-450f-8837-c1be54add358 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbda4d70-8437-4dbf-88f1-2288d9cf30a6 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Data quality in imitation learn- ing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 021b381e-610e-4990-8bc2-80032795aa18 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Conservative q- learning for offline reinforcement learning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e1add7f7-cee4-4bea-a5b0-843957fda167 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Learning to summarize from human feedback,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation dd6c7b63-915f-4d85-870b-91ffaea0f988 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Deep reinforcement learning from human preferences,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dde8dfa-2086-414e-b676-2deb21d13a71 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Stabilizing transformers for reinforcement learning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6b6372a7-f651-4083-b671-057d8223725a · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning What matters for on-policy deep actor-critic methods? a large-scale study,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f09bf7b5-5e32-4408-bdd7-9f6b7ea461d6 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Training Larger Networks for Deep Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a8e40d1-c93c-42e1-8e4f-ea9e5428aebd · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5aec9e-e848-48b4-adce-5ba4702c699e · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ead807a-a00f-483b-98d3-695f3f7bd88f · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5feda6e-87cd-4782-8d98-e3d621ab3b80 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Minedojo: Building open- ended embodied agents with internet-scale knowledge,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5079cf3-34dc-4d7f-bafa-83893cb2a5af · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Language Reward Modulation for Pretraining Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641188bd-4d7f-4b84-9886-1b045d375180 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Learning to Model the World with Language
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3e71a8-1df3-4afe-acf9-ce56e2b88004 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Grounding language to entities and dynamics for generalization in reinforcement learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4c2d6656-e7ad-4a14-b2fb-bbd4988ff49e · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a93d13f5-ad75-4fea-b855-4df3adffdde8 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning PaLM-E: An Embodied Multimodal Language Model
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8fd803f-e8e3-4706-ab09-2e1cf7883c5a · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning DoReMi: Grounding Language Model by Detecting and Recovering from Plan-Execution Misalignment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b3d59818-6969-48fb-b7b8-985ba977473d · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Collaborating with language models for embodied reasoning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec3def7f-3dfb-469e-b933-b7fb45cf9005 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Prompt a Robot to Walk with Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff6a15b-a345-4b56-9074-18853aa64dc7 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Code as Policies: Language Model Programs for Embodied Control
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 711876c2-32e3-47eb-ad4c-4e7bfbd5ee24 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Vision-Language Models Provide Promptable Representations for Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca8f1f87-fbd0-4df1-a885-97eb384485d3 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Learning to summarize with human feedback,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6956eb4e-c778-4d35-9b66-da394f6e2173 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a124d586-bb0c-40ce-bc27-b9262472fcf5 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7674b840-e200-4f01-95af-002683741b02 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Grounding large language models in interactive environ- ments with online reinforcement learning,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dcba127-0551-4671-8812-e3dc282d1bec · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Large language models as generalizable policies for embodied tasks,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6af1a7e6-4817-452c-8227-94dc0a889695 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 206a948f-f3e9-49d2-bbf5-db00b6626942 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Learning transferable visual models from natural language supervision,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 979533f5-be29-4925-b3a6-78d0b5736011 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9a4171-2f6c-44fd-9d3d-0d65787c0651 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Instructblip: Towards general-purpose vision- language models with instruction tuning,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 53383857-eaf3-44ae-991d-af29289cb0ac · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Set transformer: A framework for attention-based permutation-invariant neural networks,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cd51998c-a8aa-40ec-80e4-9a2de31f642a · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Multi layer perceptron,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 143775d4-b6a4-4799-9166-1102c6e1380b · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning LoRA: Low-Rank Adaptation of Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93fe08f0-e7bb-4869-b9eb-f37508278cea · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning TAIL: Task-specific Adapters for Imitation Learning with Large Pretrained Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ad8d96-bdee-49cf-abbb-fd04e117f5bc · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning RoboCat: A Self-Improving Generalist Agent for Robotic Manipulation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95432a34-162e-4957-92e8-8ef20c55a993 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Catastrophic interference in connec- tionist networks: The sequential learning problem,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff449bde-51f1-4d24-9e9b-17008d405a6d · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Octo: An open-source generalist robot policy,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8a1dd051-fe4d-4618-9beb-509c30ec7095 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e395ff8-c048-45d3-9100-51c08801f75f · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af934db7-e170-40db-9e11-6854753de815 · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8df84912-ce53-425a-8214-f4e6b0672e6a · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 093aedf0-9136-4e86-a557-d2c2d2d4afca · outbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4dcedf2-a314-48ff-a734-3fb524645888 · inbound
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b0f97b08-27e5-4fcf-a420-409b0c1ef202 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8d2ecca4-10a1-4fb7-ab40-3303c1c91845 · inbound
Generative AI Act II: Test Time Scaling Drives Cognition Engineering Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed18330e-329b-432b-bc7a-086823ea5350 · inbound
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7108ed8c-d69a-4c54-b1fd-20506952726b · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a27cf8f7-cea0-43b4-b7c8-1ec0799ea99a · inbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c594d43-1aa7-4147-8322-d05287c928d0 · inbound
Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65301ea4-71ca-451a-a691-056ceb3600ae · inbound
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 689c3033-cf16-4782-a3a7-eb31f2b0ae21 · inbound
SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2634ff7d-e50f-4d87-b330-b96d3689fa0c · inbound
Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd37d08d-cc3e-430e-9226-c59979728f02 · inbound
Leveraging OS-Level Primitives for Robotic Action Management Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35ba3196-ca77-4c1e-8e83-1549cb75e2ac · inbound
Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2725aca8-5c7c-48f3-8fbf-57cd903c3e8f · inbound
Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfe0573-d8ee-4686-9445-2942dda656c4 · inbound
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1ab1b8f3-bc2c-4628-bbb2-d0007fb19793 · inbound
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b2c27e-2eaa-4d5d-bc6d-f8c540b37840 · inbound
Ctrl-World: A Controllable Generative World Model for Robot Manipulation Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 90b52234-e401-4d30-be29-95df048fed96 · inbound
Reflection-Based Task Adaptation for Self-Improving VLA Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 96f1fb71-e18f-481e-972d-a1fafab88cf6 · inbound
$\pi^{*}_{0.6}$: a VLA That Learns From Experience Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1ab87388-22a1-49df-b0e9-e41fa555e3f4 · inbound
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50f1e294-9b47-479e-a108-da46695ada15 · inbound
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58ea4866-2491-401c-8be1-2517f1eb356e · inbound
VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a8596dfc-c014-4905-aa7b-adc13a2adb55 · inbound
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0d17fbb1-bf9e-414c-b44b-a41290644b48 · inbound
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 694260c5-fe28-4f0e-baf9-87fcd39c434c · inbound
PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bbc803a6-8cfe-467b-9d40-981a1f8da0fa · inbound
Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation? Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f1ace14f-8eea-428c-b697-830e5a049e2a · inbound
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 44d8d85a-3225-42b0-bb30-1eb410e04121 · inbound
Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 56c18ea0-13ad-4965-98ff-870fb0e1512d · inbound
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 133
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 510528f5-e4b1-4196-8011-37b2e07639e6 · inbound
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bbde08bf-32e7-433b-9ce7-afe9dc3ef770 · inbound
AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 69d21ad9-a07a-4301-af0e-5d08773fa56a · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5a0155eb-bdf8-4494-bf6d-5b47aabf7ef9 · inbound
Adapting Generalist Robot Policies with Semantic Reinforcement Learning Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 844fe8d3-167e-4ef6-a8df-f31242600c73 · inbound
RL Bootstrapping of OpenVLA-OFT for a Novel Robot Embodiment Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edfee9b4-2bab-447b-8495-e59005959870 · inbound
TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.