Pith. sign in

Paper Citation Record · LEDGER

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

As of 11 August 2026, this Paper Citation Record lists 91 of 91 outbound references and 2 inbound Pith citation observations for arXiv:2602.12628.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.12628 v4

Coverage vector

measured 91 of 91 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T23:49:05.337733Z

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:34:04.267745Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T10:48:02.217286Z

Reference resolution

91 of 91 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bd3e8958-b588-4be5-9e15-d7b15b02246c · outbound

This paper cites Learning dexterous in-hand manipula- tion.The International Journal of Robotics Research, 39 (1):3–20, 2020.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Learning dexterous in-hand manipula- tion.The International Journal of Robotics Research, 39 (1):3–20, 2020

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.118801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.118801Z digest=sha256:a69aff5df97beabd1fb100942d1b76e2e270181e7985479aad63e72c2effc8c3

Observation b96170df-287b-48cd-8f92-0872773422da · outbound

This paper cites From imitation to refinement- residual rl for precise assembly.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models From imitation to refinement- residual rl for precise assembly

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.210423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.210423Z digest=sha256:e0ce6644e37a6f5fb811d8d2c06410c82485609e6438d97cb666a90d718463f0

Observation 69227d5a-2832-4df6-8289-6c8aec76d1bb · outbound

This paper cites Qwen2.5-VL Technical Report.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.297050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.297050Z digest=sha256:a475ae19c348ce31cf8f90d75a5f8e76842f18722ae8cecaac6c491e4fca7a00

Observation 3d90655f-407e-4843-982d-9260ac109a8d · outbound

This paper cites Efficient online reinforcement learning with offline data.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Efficient online reinforcement learning with offline data

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.369035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.369035Z digest=sha256:53cb2738475bd9eacedc7194e2c19d88cfe939bf9a05ea8e7744d3ada41cbae6

Observation c7cf64e9-b3fa-4ef7-8f2a-60b138c50d0e · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models PaliGemma: A versatile 3B VLM for transfer

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.501648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.501648Z digest=sha256:6bbb27de1c36dcd2b90c85a92478a2ca514b91d86f0d37e90b8ac39fc1847fcd

Observation 6afe496c-9f34-457f-b06e-d5df297bab06 · outbound

This paper cites The r2r framework: Publishing and discovering mappings on the web.COLD, 665:97–108, 2010.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models The r2r framework: Publishing and discovering mappings on the web.COLD, 665:97–108, 2010

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.607861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.607861Z digest=sha256:12537bb357936bb53e3b0b9059e957d11fc48e61577652bea13ce98a29b750c8

Observation 2a0f81bc-9fbc-4c27-990e-02e94bea70eb · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.703806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.703806Z digest=sha256:34e514b1e58a46b0190fd92eb4b9b754f1414cc56fa37280891faa1de9c68033

Observation 8a2f0396-42a1-447f-b8f0-11a5f8c085be · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.768219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.768219Z digest=sha256:8c6e88d5a27db93f4e7cc04639dcf535c05a402fe1eb7db44bb96eec1938661e

Observation a1c86820-8c0d-4a06-b185-99460bf02c71 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.812866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.812866Z digest=sha256:5a1ea6b4fc837c9634336f64d70f16daa7b28fab9bbfe6af3faf91abb8dad989

Observation 6ff5d23c-c347-48d9-bd5f-6dba6bfcc7eb · outbound

This paper cites The ycb object and model set: Towards common benchmarks for manipulation research.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models The ycb object and model set: Towards common benchmarks for manipulation research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.871238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.871238Z digest=sha256:6bbe4e24327f00beccbe42a0559edc04a74662b44cb76dd5faff152a845a2d5b

Observation a286c1d0-5f10-4073-8552-e4318edba703 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models ShapeNet: An Information-Rich 3D Model Repository

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:55.960759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:55.960759Z digest=sha256:8f48f0452eb2157d15b8496df4f1d63cb02ca8d10d838bae6f44c0df3c62560c

Observation b797cd48-dee6-449b-8469-a669a19acbf1 · outbound

This paper cites Conceptual 12M: Pushing Web-Scale Image-Text Pre-Training To Recognize Long-Tail Visual Concepts.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Conceptual 12M: Pushing Web-Scale Image-Text Pre-Training To Recognize Long-Tail Visual Concepts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.083452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.083452Z digest=sha256:9552ab988ce55a9cab68cddd01747a0122a088bbd072d32e8a983c0ac5fa7c21

Observation 837802a2-2908-4492-b2d1-0909ea1a6b99 · outbound

This paper cites Closing the sim-to-real loop: Adapting simulation randomization with real world experience.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Closing the sim-to-real loop: Adapting simulation randomization with real world experience

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.193968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.193968Z digest=sha256:180d2f666493c2d3de325569cd363eb778023688e7c363ae65addbbd56597001

Observation 282fd24b-90f3-4a70-a29f-40dfbc078c62 · outbound

This paper cites URL https: //arxiv.org/abs/2510.25889.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models URL https: //arxiv.org/abs/2510.25889

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.294254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.294254Z digest=sha256:2da60939f409d6fbee480be9e441d0cb2aa0def8d96208a026df8d5d14d4e038

Observation 5c970ca3-6b6f-49ae-8cda-a266466aad89 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.406404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.406404Z digest=sha256:50d54d06791d95ecccecfbfc183b1c622f0ffc44b08b4a180b6ba1d6ee288815

Observation 9de5d3ad-0c5a-4c49-aefc-af1f6632ab35 · outbound

This paper cites RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.520705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.520705Z digest=sha256:ea576e419b458b358075b28970ec90d7ee4f116a49b3de0fff903152915bf7ef

Observation f1e56d64-6c93-495d-ac3f-89a95ad83e58 · outbound

This paper cites Generalizable domain adaptation for sim-and-real policy co-training.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Generalizable domain adaptation for sim-and-real policy co-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.633307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.633307Z digest=sha256:532f46b1153f580dee06f65334fc0d4bc3383633b2f0f36cafefa8cb916ce69b

Observation 0e448b06-eb78-4a2c-8470-4629ebbf93a5 · outbound

This paper cites Automated Creation of Digital Cousins for Robust Policy Learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Automated Creation of Digital Cousins for Robust Policy Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.712709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.712709Z digest=sha256:c3af1cae0610b679097b82b60745d07e0b0dcdc5951c04f2a98dd1183475d37d

Observation b10b93c9-fef6-4367-812e-795dbc4cdb4a · outbound

This paper cites Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799– 35813, 2023.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Objaverse-xl: A universe of 10m+ 3d objects.Advances in Neural Information Processing Systems, 36:35799– 35813, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.824247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.824247Z digest=sha256:6e662d0070c695598272208b284d0164e1f6517e7064cb7400d57762c63253ae

Observation 0dfdf214-0518-40d1-bf97-fd743dcbda24 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Objaverse: A universe of annotated 3d objects

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:56.932854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:56.932854Z digest=sha256:5682883ca0dec6acc0cbe3140a9e25ba97cdd970d93a8d9296b72b66a1941bae

Observation 4aa962d1-a4cb-44bf-a64d-7b077b1f8baf · outbound

This paper cites Challenges of Real-World Reinforcement Learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Challenges of Real-World Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.035363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.035363Z digest=sha256:143f2ae2ed3a224c4350df475d698df5aecf0e9c140ab55bec539480a9e2ae22

Observation 8d11e19f-87e8-4781-9fdf-7d0fe89a0f52 · outbound

This paper cites Bridge data: Boosting generalization of robotic skills with cross- domain datasets, 2021.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Bridge data: Boosting generalization of robotic skills with cross- domain datasets, 2021

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.109819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.109819Z digest=sha256:d727444d72d5a86c153d4838c31ecaa5fa961feb8ff854f7d14c224f81dc165e

Observation c740db79-c3d8-4792-bc17-2e305cd322bb · outbound

This paper cites Sim-and-human co-training for data- efficient and generalizable robotic manipulation, 2026.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Sim-and-human co-training for data- efficient and generalizable robotic manipulation, 2026

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.156984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.156984Z digest=sha256:011c217a1ac74a533b3a1ea063d9ac91479b22bbf3caba0eb6b2383eaab73899

Observation 5d9f1f76-c22e-42fc-b644-5dd28afaf6ba · outbound

This paper cites ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.271382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.271382Z digest=sha256:2be9a85cfe80e137c685429c97dfd804c544eef0621911eee7e9abcd7d13b95b

Observation 5edaf50e-4315-45b6-9ea9-9009fb0c46f4 · outbound

This paper cites Airbert: In-domain pretraining for vision-and-language navigation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Airbert: In-domain pretraining for vision-and-language navigation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.427603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.427603Z digest=sha256:f81ca2cc5116e466e31e13fda9003d47b76e7633e5d0545df3ae11f5cd2f6faa

Observation 108bd3a8-8724-4870-81a5-7b083f5f2c9c · outbound

This paper cites Towards learning a generic agent for vision-and-language navigation via pre-training.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Towards learning a generic agent for vision-and-language navigation via pre-training

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.564187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.564187Z digest=sha256:480c4a83fc9500fb7f2a52cf668488fe1cc134008334edb1cc2d942ae07b104e

Observation d9b80694-0e5c-44c1-8c72-341f18fed864 · outbound

This paper cites Vln bert: A recurrent vision- and-language bert for navigation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Vln bert: A recurrent vision- and-language bert for navigation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.715735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.715735Z digest=sha256:6c30cb6a77657798b8220e9ee38eeb9b311c6070bf6b91fb978e76b1d82fa4da

Observation 239b05dd-2c7f-4c2e-b40d-47c5b4b1c763 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.810846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.810846Z digest=sha256:e8b8ec752740fbbb8451091cd0c3d9958d006bcc6ab9657a2d7daf97cb564fbe

Observation 65c3e564-9f56-4902-ac84-9ac87c0dd074 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:57.949446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:57.949446Z digest=sha256:e082ce63c0579d61aafecb6bddc7ff5a1768c67ebc2b71e01b7a1739deb2867a

Observation dfe368ce-6ede-4d57-a0e1-890224b89277 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.046107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.046107Z digest=sha256:5a4e348d6cf96e9bf26b6d73ee8f2fb19de9bd90e808a7d38e81d231f60052d0

Observation 7a51e6b2-0f00-43e3-9c36-19347d99f9c8 · outbound

This paper cites VIMA: General Robot Manipulation with Multimodal Prompts.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models VIMA: General Robot Manipulation with Multimodal Prompts

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.160509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.160509Z digest=sha256:abe71c0ecfb991ea2f01dadb2cc1135e2b4140f72d0469a785dc5e05973186a1

Observation 0bf2d53a-8726-4b2a-a61d-2a2d6bd29683 · outbound

This paper cites QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.328749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.328749Z digest=sha256:3bf6fa7a1cc0830ef0a55bb174eb4e870b2d5f24f60bf266389743bfe9931a5b

Observation e09bd3ca-cfd4-4820-b259-346a70babba8 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Trans.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 3d gaussian splatting for real-time radiance field rendering.ACM Trans

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.408204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.408204Z digest=sha256:a6b63ea5bb5ed44478de44fb766f32b4953aef56bc672c4648f453ec21ba2dea

Observation 6d2b50bc-b2b8-407f-aed5-aa3a65b80ec7 · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.485677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.485677Z digest=sha256:00d08eac7b5b581b98e2e8a297cbb3c55b0194a372e5b5dd1d884534a0066261

Observation 50fed8f2-38a4-4fa0-9fed-1b8d4fb33790 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.559248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.559248Z digest=sha256:e5e5a684ead7f3f51ebabd1b0cb5871f85ff6a077a7355c5f8f6f2a787b1aaf1

Observation a424bdda-7842-42f6-be7f-35986eaea3ff · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.670284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.670284Z digest=sha256:27f697ffcc3d466363a83a978fdc5bfd2c44273d1fecfcf62eb36c51dd7c017d

Observation bedf0b36-102c-48fe-a757-0962240c6c86 · outbound

This paper cites Room-Across-Room: Multilingual Vision-and-Language Navigation with Dense Spatiotemporal Grounding.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Room-Across-Room: Multilingual Vision-and-Language Navigation with Dense Spatiotemporal Grounding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.731930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.731930Z digest=sha256:ab59e6c5ee3121b726cbfb2ccbae080e12a4a0cd9030ab9e7eb610f67a9da823

Observation 05ecb8a6-185b-4756-b3c0-6f57b7d934ed · outbound

This paper cites SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.812303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.812303Z digest=sha256:c36429c36426179a2f15d49a66c71eed5f2e06738437a0796596bfba999f4440

Observation 60167d4b-159f-40da-ac14-90a2e039470e · outbound

This paper cites RoboGSim: A Real2Sim2Real Robotic Gaussian Splatting Simulator.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models RoboGSim: A Real2Sim2Real Robotic Gaussian Splatting Simulator

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.850365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.850365Z digest=sha256:5a557dd64ceab52db61718c40a3a44eab281f4fc7faf68ccc48cef3add37a3fa

Observation 836d58d0-5f70-4237-898a-4f4496fc8a6d · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.arXiv preprint arXiv:2512.01801, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.arXiv preprint arXiv:2512.01801, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.907946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.907946Z digest=sha256:61542977cfae4516c55293b74b7f6d332c7ebf0d81bc0467cf763a1a8692a37e

Observation a91bec37-e355-4a20-8134-4371d87e530d · outbound

This paper cites Flow Matching for Generative Modeling.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Flow Matching for Generative Modeling

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.051315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.051315Z digest=sha256:396a1bb527eea6ee0d76b58a793f55b09d70b11fd94bd5042f1ac41bc582f83c

Observation 71584382-a73d-4a81-bcaa-242864dff902 · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Flow-GRPO: Training Flow Matching Models via Online RL

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.184987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.184987Z digest=sha256:7f54a532aaf7c4b538245c6c23d18ee2e631f4ea9c3f2590c3e737c69c01cdb8

Observation 89996bc0-322a-4867-945e-fca94deafa95 · outbound

This paper cites What can rl bring to vla generalization? an empirical study.arXiv preprint arXiv:2505.19789, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models What can rl bring to vla generalization? an empirical study.arXiv preprint arXiv:2505.19789, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.240329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.240329Z digest=sha256:15838e2e0deb5d4adc20628d8554b108c3e1c3bff69528912ae2e86a2ab5f6e9

Observation 932f9bad-285c-41c0-abb1-b8b2dc343239 · outbound

This paper cites Decoupled Weight Decay Regularization.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Decoupled Weight Decay Regularization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.285658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.285658Z digest=sha256:3bbe1c6684f70525e9fbe0a4529277ad96f962de31a6edceca22350ec2e6ab24

Observation e5a2bece-4eb7-4d2a-a9d0-23785dfc9c4c · outbound

This paper cites Serl: A software suite for sample-efficient robotic reinforcement learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Serl: A software suite for sample-efficient robotic reinforcement learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.379112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.379112Z digest=sha256:9f877612a14f3417672c0219a88c4147d0b5350f27bb15badd204941a8e6034b

Observation df251004-bb0e-4d1d-a5a7-ebe2c30610a1 · outbound

This paper cites Sim-and-Real Co-Training: A Simple Recipe for Vision-Based Robotic Manipulation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Sim-and-Real Co-Training: A Simple Recipe for Vision-Based Robotic Manipulation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.417592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.417592Z digest=sha256:75ad45daa41cda73f420b74177d114190d4bf366692781fdd8e0e8cc7d633116

Observation 950468ad-db31-4a23-a1e5-7554bd7586cf · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.505472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.505472Z digest=sha256:5043dc1c84f970879238dd256ba30b7f5f50c41adf5e8fddebba64dcfe431fc7

Observation a2609a87-753c-4d44-aea1-304a1c8f67c6 · outbound

This paper cites MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.546792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.546792Z digest=sha256:dc2bd40016cb087f5f74857c3d23f21f88b640e11c04ee93fa60fc87a6b4efd8

Observation 2fa6c766-59b2-4f47-b085-d21981d3ee62 · outbound

This paper cites Active domain randomiza- tion.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Active domain randomiza- tion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.610401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.610401Z digest=sha256:036c581b20e181f9ebf2304235be795900e7b321410f2208cd1cf031af927528

Observation c6ac8c67-9a24-42ea-8a07-350fc3638935 · outbound

This paper cites ManiSkill: Generalizable Manipulation Skill Benchmark with Large-Scale Demonstrations.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models ManiSkill: Generalizable Manipulation Skill Benchmark with Large-Scale Demonstrations

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.699822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.699822Z digest=sha256:09f68d3f73b30006037703f771fd3ea980b674eb2a9a5f9fd0c3317b8296b6f5

Observation 703261df-11e0-4c25-86a5-e76eeca68261 · outbound

This paper cites RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.774612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.774612Z digest=sha256:5cfc040fb96ab05e63e5c23e3883dcb4a762e1b545e8a291a3f77755a6d08738

Observation 25ac4953-4123-49f5-917a-71546725ce92 · outbound

This paper cites An algo- rithmic perspective on imitation learning.Foundations and Trends® in Robotics, 7(1-2):1–179, 2018.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models An algo- rithmic perspective on imitation learning.Foundations and Trends® in Robotics, 7(1-2):1–179, 2018

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.867282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.867282Z digest=sha256:3bf03d330804df5a1139357dca4195b0b4293fd1474680b6107a5808048806b9

Observation df2bc198-d76b-4299-afc0-1e701b94ac0d · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:59.984155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:59.984155Z digest=sha256:4ceba702f35c4aa6fd238cc4646dcfb7d3fb840d0ef7d358428261fe4021e8aa

Observation 7c3c28f8-d0e4-417d-b2ad-ec32ed6b190b · outbound

This paper cites Sim-to-real transfer of robotic control with dynamics randomization.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Sim-to-real transfer of robotic control with dynamics randomization

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.162601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.162601Z digest=sha256:7609861e2dceb59504b5756accf84b0b6274722820e46d1b6e345551d7f40af6

Observation aa96b96c-8ff1-400e-b1ae-76e729852717 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models A reduction of imitation learning and structured prediction to no-regret online learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.240540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.240540Z digest=sha256:763290814e16abc9880414d02dbe26ffcee0b940a3681a0545d19c7c138fa320

Observation 5a974846-25a1-44bc-8ed3-3ed552818365 · outbound

This paper cites Habitat: A platform for embodied ai research.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Habitat: A platform for embodied ai research

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.315846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.315846Z digest=sha256:dde3679f7d38fe3625e5a695ba27ea9b502b43bcd8f0d208cd2ba3c5d01ac150

Observation ce9a3e1d-2efd-4116-8ed2-deabb6002dce · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278– 25294, 2022.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Laion-5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278– 25294, 2022

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.402882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.402882Z digest=sha256:155607544fb033d823a856cc80267694d8b34fc1f247423cb409e758027650c9

Observation c1b56d95-8ea5-4dc1-b9b8-9ff0b612fee9 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Proximal Policy Optimization Algorithms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.537756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.537756Z digest=sha256:a28a4174871dafe7f0a877199c866e7b4c8b8350084630a0483c15e6a22e3373

Observation d78e8f76-18b2-45fe-93fc-a85e90ee7dbc · outbound

This paper cites Videovla: Video generators can be generalizable robot manipulators.arXiv preprint arXiv:2512.06963, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Videovla: Video generators can be generalizable robot manipulators.arXiv preprint arXiv:2512.06963, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.617702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.617702Z digest=sha256:3bbc25ad386dc18af3a6f54d8b8f2967704e5b5d57d35eb1c9d7164685c7aa66

Observation 518640b5-0678-4224-bfef-de014e116198 · outbound

This paper cites ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.760526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.760526Z digest=sha256:427a3b138c7cb22f371c93425fd2a6584c5615318104b0b22ad4a5692c651c24

Observation 2929f278-f9e9-4395-a879-a62502d26922 · outbound

This paper cites Evaluating gemini robotics policies in a veo world simulator.arXiv preprint arXiv:2512.10675, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Evaluating gemini robotics policies in a veo world simulator.arXiv preprint arXiv:2512.10675, 2025

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:00.934757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:00.934757Z digest=sha256:3cfd9113a65ccb0f460cb8e4edf91ba4468eb2111b45a3c9ab5ec7263101e85a

Observation 31c2bb92-1ba1-4463-8526-836c720c865c · outbound

This paper cites Gemma 3 Technical Report.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Gemma 3 Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.043861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.043861Z digest=sha256:c527492f9a4907d2534087cf12d2422b7ecc49ef6da7231850097fa3f7074651

Observation 4dcba3d1-350b-46a4-8cca-5fa188a81264 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Octo: An Open-Source Generalist Robot Policy

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.157381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.157381Z digest=sha256:a2061367b4343cfa875804771f6bd911738eafaaccf46289902089b416fd7869

Observation 1f4b8432-d611-4dce-a9e7-cf0021c28436 · outbound

This paper cites Vision-and-dialog navigation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Vision-and-dialog navigation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.265440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.265440Z digest=sha256:cd219c459782d188ac070493ff9fa66b1534190fcac00d13b87b75a7a92ec6ac

Observation 60086f30-49d5-494b-b9d9-3bd2917af922 · outbound

This paper cites Domain ran- domization for transferring deep neural networks from simulation to the real world.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Domain ran- domization for transferring deep neural networks from simulation to the real world

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.460377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.460377Z digest=sha256:bb912f2a90db2129517d0ec378797815c6c299978f11a9b2311969b7823700aa

Observation f102ba7b-60ae-49e8-9674-382b5881c67c · outbound

This paper cites Mujoco: A physics engine for model-based control.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Mujoco: A physics engine for model-based control

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.642862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.642862Z digest=sha256:436ada1594aa5b3edbdef6102267be5d8488b878355ce08b3616b697a3f4ac86

Observation 749f911e-6f1b-4ff4-9468-3a017fb282c6 · outbound

This paper cites Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.747074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.747074Z digest=sha256:b24a3f44ee2931a79a7564f885620c5d5f4823a3c8634f041e865bb55ef5f5d1

Observation 06268b50-dc20-40d1-a435-083977e4b91e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.861640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.861640Z digest=sha256:df39a20e6be5a4c0d5c8a0ed696af66c59d4038843410b522e7242393ad7ecae

Observation eecfe2d6-e0f6-46ca-96b5-60bcc34f62cd · outbound

This paper cites Bridgedata v2: A dataset for robot learning at scale.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Bridgedata v2: A dataset for robot learning at scale

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:01.996145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:01.996145Z digest=sha256:d48dd71d88a8cf2cc8438137c5de7a21333153a6b66fb8f081c0b6322c9e1186

Observation adf8a657-1748-4a61-b9fc-ebfe8c7f001c · outbound

This paper cites Empirical Analysis of Sim-and-Real Cotraining of Diffusion Policies for Planar Pushing from Pixels.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Empirical Analysis of Sim-and-Real Cotraining of Diffusion Policies for Planar Pushing from Pixels

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.144147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.144147Z digest=sha256:b70bc06f0c87a921b178d7d43830abfd13d0b411ba543765ce022570411e1bd5

Observation a4dcdf56-bdfe-412a-afb2-6fc6fa64e8ab · outbound

This paper cites Rl-gsbridge: 3d gaussian splatting based real2sim2real method for robotic manipulation learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Rl-gsbridge: 3d gaussian splatting based real2sim2real method for robotic manipulation learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.285758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.285758Z digest=sha256:5d4d1c51ae03c3380339a53e0d17a192ef74aa5d7c66ec1c0a397374564b7b49

Observation b347b7a0-e63b-45f8-ad9b-16c7a1d3732f · outbound

This paper cites Qwen3 Technical Report.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Qwen3 Technical Report

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.431397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.431397Z digest=sha256:a779833e2e5b5b2aa9c8fc790fb1bfd10446cd1869c10a48c7e7997697693919

Observation a285db93-4cd4-4295-b0f8-98a0b2c544fd · outbound

This paper cites Invari- ance co-training for robot visual generalization.arXiv preprint arXiv:2512.05230, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Invari- ance co-training for robot visual generalization.arXiv preprint arXiv:2512.05230, 2025

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.592585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.592585Z digest=sha256:e33f736bb7eae61a7d8be92a7553bdb6f39f747183959f17fbf7ef5972853e4b

Observation dbf81d85-2fb7-4d0b-bece-c332fd77194b · outbound

This paper cites INeRF: Inverting Neural Radiance Fields for Pose Estimation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models INeRF: Inverting Neural Radiance Fields for Pose Estimation

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.818761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.818761Z digest=sha256:d485a30cd55ce58dbea1b75222aa95e9490f209d90160a7572b6c0f4541e0c45

Observation fb0b1bd3-feca-4835-b7d4-f03179d5f686 · outbound

This paper cites Natural Language Can Help Bridge the Sim2Real Gap.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Natural Language Can Help Bridge the Sim2Real Gap

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:02.985771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:02.985771Z digest=sha256:ba250795944cd8c11a408c074ba3a566aefc485fa1e5fdefb04e20d297672a51

Observation 3031cebb-d6a7-4e25-95f0-9f8a73b6d9e4 · outbound

This paper cites Rlinf: Flexible and efficient large- scale reinforcement learning via macro-to-micro flow transformation.arXiv preprint arXiv:2509.15965, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Rlinf: Flexible and efficient large- scale reinforcement learning via macro-to-micro flow transformation.arXiv preprint arXiv:2509.15965, 2025

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.144415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.144415Z digest=sha256:780224ca264a76b7fdd61ab410cc8abc84d765d3c758344eb47c80d09ad15cc5

Observation 1a8bc444-43a8-4f04-963c-e6bc63f7afc4 · outbound

This paper cites Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.275426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.275426Z digest=sha256:0055c8bf4e218dcff2c0dd56ca86c24777648bf813f854dfef921b580c5bc982

Observation 0600883d-39e6-473b-b61e-474abd25378b · outbound

This paper cites Rlinf-vla: A unified and efficient frame- work for reinforcement learning of vision-language- action models, 2026.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Rlinf-vla: A unified and efficient frame- work for reinforcement learning of vision-language- action models, 2026

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.416040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.416040Z digest=sha256:eb6e1f637f4854ac3d0756f741d5552a21bdf024cffec02a691b54452a7f8d93

Observation eb291e34-4f36-4747-92b3-121857de97b8 · outbound

This paper cites Real- to-sim robot policy evaluation with gaussian splatting simulation of soft-body interactions.arXiv preprint arXiv:2511.04665, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Real- to-sim robot policy evaluation with gaussian splatting simulation of soft-body interactions.arXiv preprint arXiv:2511.04665, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.549584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.549584Z digest=sha256:676ea78fe7e8b275b3aea6350354a65342003ff2109b491426d93fe16da7f307

Observation 4cd4cb20-7d0a-4d33-833a-a0527cd597af · outbound

This paper cites VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.695895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.695895Z digest=sha256:072391c38d2dc430cd238ccd838be90fa9bcddfe82791f95a0e0e9cc169d2890

Observation 07c31fe2-9dc2-4553-8702-8e77d312fe0b · outbound

This paper cites Re- inflow: Fine-tuning flow matching policy with online re- inforcement learning.arXiv preprint arXiv:2505.22094, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Re- inflow: Fine-tuning flow matching policy with online re- inforcement learning.arXiv preprint arXiv:2505.22094, 2025

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:03.896128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:03.896128Z digest=sha256:35227e14af7b8fe1adc82f20d443b6fa5b1cca92809f78ff82bbd6f9d641ae0f

Observation 00b74cbe-ac69-48ee-b69e-de9c97a6d677 · outbound

This paper cites Sac flow: Sample-efficient reinforce- ment learning of flow-based policies via velocity- reparameterized sequential modeling.arXiv preprint arXiv:2509.25756, 2025.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Sac flow: Sample-efficient reinforce- ment learning of flow-based policies via velocity- reparameterized sequential modeling.arXiv preprint arXiv:2509.25756, 2025

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.008280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.008280Z digest=sha256:4c4217284151527b7ac8b1327d02b69edd52462779035cb4c183ef41871f9724

Observation c1931af5-89c4-4f74-88ec-e0c84efd5bb7 · outbound

This paper cites Learning 3D Persistent Embodied World Models.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Learning 3D Persistent Embodied World Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.208684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.208684Z digest=sha256:4b1e2ea4de3e677033162340ea5527565e7f2732102ef754a836fea568beb573

Observation 246a5e94-f1fa-46c3-83f7-fa6625419b92 · outbound

This paper cites IRASim: A Fine-Grained World Model for Robot Manipulation.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models IRASim: A Fine-Grained World Model for Robot Manipulation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.287251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.287251Z digest=sha256:78c29530280f3d1220570ce1063393631b383b4f318849ece1a45f7447c89423

Observation 16cff3e9-b458-4c63-a547-38bf26727c70 · outbound

This paper cites Rt-2: Vision-language- action models transfer web knowledge to robotic control.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Rt-2: Vision-language- action models transfer web knowledge to robotic control

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.343005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.343005Z digest=sha256:83b6e0455c08b6c9dfa695b1c122cd16587120716243d7f57f87748320dd99b0

Observation ca0503c5-21c1-4943-93ed-a6f1284dcd1d · outbound

This paper cites 9 visualizes the four table-top manipulation tasks evaluated in our experiments.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 9 visualizes the four table-top manipulation tasks evaluated in our experiments

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.516794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.516794Z digest=sha256:fbca2754fb4daaf2175b1d4f2bd9af3f1da28662a5d084677724302c11f9cf5b

Observation 0beb48bd-0b27-4810-844e-cdddc2e930a9 · outbound

This paper cites 10 shows the manipulated objects used in both simulation and real-world experiments.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 10 shows the manipulated objects used in both simulation and real-world experiments

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.656900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.656900Z digest=sha256:ef831098d5f2cc63e7f53868b27d358d41278291be15fc15a0adfcb4fa9122d4

Observation 3c471d34-af63-4afe-9dd1-1e4b02d7b832 · outbound

This paper cites 11 illustrates the randomized regions for the four real-world tasks.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 11 illustrates the randomized regions for the four real-world tasks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:04.834261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:04.834261Z digest=sha256:ddee3311d260e5ed4e9013c074afc3c78c2925e0aae277b41c2e051070d46bfc

Observation 8f37aa55-ad2d-429d-abed-49c92c8b78d3 · outbound

This paper cites an unresolved cited work.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Unresolved cited work

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:05.027202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:05.027202Z digest=sha256:323ac0398165625f3e789d44614a0ab67eba59bda26e14bc1e8de2f1d3063c99

Observation a40a66c7-a95e-4ee7-9b81-ab32c117a8f0 · outbound

This paper cites an unresolved cited work.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:05.208709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:05.208709Z digest=sha256:862fa5f53753b6cf8d974314c6c3564496496304d0305f1097437e9a69b57d15

Observation 84e0a516-b098-40c7-9a6f-479f3fa45090 · outbound

This paper cites 12 shows the success rates of different models on each task in the simulation environment during RL training.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 12 shows the success rates of different models on each task in the simulation environment during RL training

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-02T23:49:05.337733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:49:05.337733Z digest=sha256:e9c7951b376b0e42b4f5ad6af634aef398e7a1b940ccc9619e0fb9eb27ebcef7

Pith citing papers

Observation 1c1286bd-3121-4389-8f5f-db24d60416a9 · inbound

TacCoRL: Integrating Tactile Feedback into VLA via Simulation cites this paper.

TacCoRL: Integrating Tactile Feedback into VLA via Simulation Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:48:02.218471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T09:51:56.563479Z digest=sha256:d04c9ca86a4952b0008f51846d55f3703df6a9f032395a661897ccf0df697ca2

Observation 592d2881-d740-4607-8194-2e85d338ba83 · inbound

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning cites this paper.

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T03:34:04.267745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:34:04.267745Z digest=sha256:bb8f3b3fcb57a43c8695e8630b46130cf058db5cd3b5adf5ad7dcc0dc62828c3