Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:36:20.594100Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 5 inbound Pith citation observations for arXiv:2602.20220.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:36:20.594100Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T00:54:12.099045Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T19:30:07.865484Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d0f41b2c-084d-4af3-abd4-5bde4a65733e · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Sutton and Andrew G
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 199c9738-505b-4318-b676-c2f9038dd93e · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Chapman & Hall/CRC Artificial Intelligence and Robotics Series
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea16bc4e-da9f-45d8-94a4-48aa0f8cc747 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Learning agile and dynamic motor skills for legged robots.Science Robotics, 2019.(Cited on pages 1, 4, and 16)
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4db4aa6-7aa6-472e-8232-0ea10948f984 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots IndustReal: Transferring Contact-Rich Assembly Tasks from Simulation to Reality
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b539d2b-dbe1-4d68-b143-319871fdcb19 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Diffusion policy: Visuomotor policy learning via action diffusion.The International Journal of Robotics Research, 2025.(Cited on page 1)
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b468079-d591-46a5-b6b9-a3df30d7e2e5 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots BeyondMimic: From Motion Tracking to Versatile Humanoid Control via Guided Diffusion
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8bc4987-e5a0-4568-9286-8af018904a9e · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5045fc0f-ad30-4576-9aa4-c0617cf410e8 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0fe0836-7a91-45e0-b36e-41a57adab712 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots MuJoCo Playground
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18d1731a-ce5b-4d04-be6f-fa854c6de330 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Addressing function approximation error in actor-critic methods
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5b7039-aa41-41da-8867-96c6d4ea6998 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Reinforcement learning in robotics: A survey.The International Journal of Robotics Research, 2013.(Cited on page 2)
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e992df3b-93cf-4d08-bb18-e23315fe83f2 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Learning to Walk in the Real World with Minimal Human Effort
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35b970f1-1a5d-4fb4-8434-4fb20438dc4d · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots COG: Connecting New Skills to Past Experience with Offline Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e663363f-be65-4594-be85-639a72cf995d · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49603da0-a967-47cc-b000-7cf3056edd41 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Daydreamer: World models for physical robot learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49bdb634-e974-4521-9030-eea7149544a2 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Finetuning offline world models in the real world
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0afc3af8-ccc3-437b-9367-d565cf119dc2 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Efficient online reinforcement learning fine-tuning need not retain offline data
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a5941a7-1329-4d9c-825f-63ffa8271895 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Precise and dexterous robotic manipulation via human-in-the-loop reinforcement learning.Science Robotics, 2025
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeb0df54-3849-4454-b1d4-7b27d093f396 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots How to train your robot with deep reinforcement learning: lessons we have learned
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35779e2c-c688-4e04-a37c-181e12850100 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Replay across experiments: A natural extension of off-policy RL
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcb15989-78d9-47cc-8d20-4a9197518654 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Rapidly adapting policies to the real-world via simulation- guided fine-tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae6870c-b5d4-46d7-9f59-7c5d484b0d63 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Legged robots that keep on learning: Fine-tuning locomotion policies in the real world
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec06c701-36de-4cb1-b092-0e80c8cf97d6 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c8dacd8-67b1-44ed-a07e-b3560b586e1c · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots John Wiley & Sons, 2014.(Cited on page 3)
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2ec8e4b-ec90-42e5-8169-8eeb7f2886a8 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Bertsekas.Dynamic Programming and Optimal Control
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95118829-be51-4362-8de9-a2b444fb34b8 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Self-improving reactive agents based on reinforcement learning, planning and teaching.Machine Learning, 1992.(Cited on page 3)
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a6cbda2-8f21-4002-a97d-9f3f62d7fdef · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Playing Atari with Deep Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9bbb50d-629c-4143-ad84-c34ed3a9a806 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Leave no trace: Learning to reset for safe and autonomous reinforcement learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae849b38-2b39-446c-952f-017029ec61da · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Autonomous reinforcement learning via subgoal curricula
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ea8fff8-18b2-415e-b6f4-9aab26499aa2 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Autonomous reinforcement learning: Formalism and bench- marking
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ef68a0d-8401-45e4-b983-2aca514542f6 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots A state-distribution matching ap- proach to non-episodic reinforcement learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6543067f-033e-4fc1-bad9-9fa9fb4a43d1 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Conservative q-learning for offline reinforcement learning, 2020.(Cited on page 3)
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 728fc3ef-3fa6-41dc-a5c6-8930da6f9094 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Mopo: Model-based offline policy optimization.Interna- tional Conference on Neural Information Processing Systems, 2020.(Cited on page 3)
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7485183-7f90-4f72-8341-4a8b32ff8772 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Offline robotic world model: Learning robotic policies without a physics simulator.arXiv preprint arXiv:2504.16680, 2025
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bdaafc0-11e2-4038-b166-0913b520896b · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b4a616a-9301-4e32-a7e2-17f787cf13d0 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Domain randomization for transferring deep neural networks from simulation to the real world, 2017.(Cited on page 4)
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09b1d3c5-c5b2-487b-abda-34881eb05015 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Proximal Policy Optimization Algorithms
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7f03011-c7d5-40d4-8012-4ae88726a5be · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Continuous control with deep reinforcement learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12607e4a-8e25-4eb0-b54c-f04c69513eaa · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Maximum a Posteriori Policy Optimisation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0e7c28-e502-4ccc-b5c7-65ea28f899d9 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b9944c4-d5cd-40ba-b0da-54d8bf9e1c1b · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Getting sac to work on a massive parallel simulator: An rl journey with off-policy algorithms.araffin.github.io, Feb 2025
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f91f8002-d474-4132-b42d-28742e810476 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Deterministic policy gradient algorithms
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e38426cc-261d-4af9-a24b-a1b26f067720 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Equivalence Between Policy Gradients and Soft Q-Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8038379-071f-4908-8ce9-0964fd7587bb · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Approximately optimal approximate reinforcement learning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99aa29d9-ae53-43fa-a639-6ff2d9fda713 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Efficient online reinforce- ment learning with offline data
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0046c12a-d261-42b2-8c9d-f213ce832d7c · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Hybrid RL: Using both offline and online data can make RL efficient
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3c04ac4-a1e0-4474-bb0e-76dad9d3473f · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Steering your generalists: Improving robotic foundation models via value guidance
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2755953c-b859-48c9-a2d3-eb29f534a7b7 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17a6ed49-e9d1-4538-9dd9-4301de990671 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots When to trust your model: Model-based policy optimization.Advances in Neural Information Processing Systems, 2019.(Cited on page 6)
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98f63cc1-5242-4820-86e2-1075fe6e3200 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Randomized Ensembled Double Q-Learning: Learning Fast Without a Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0883ae8d-da4f-4058-987b-f6103e43e0ee · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Overestimation, overfitting, and plasticity in actor-critic: the bitter lesson of reinforcement learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae73a42f-fc73-455f-8f40-635d2680f5a4 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots On actor-critic algorithms.SIAM journal on Control and Optimization, 2003.(Cited on page 6)
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f97fd0ea-306a-40e3-b65a-d84a07841582 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Convergence rate of linear two-time-scale stochastic approximation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8803372a-5837-4c18-81e8-79baf7ef0c02 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Springer.(Cited on page 6)
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 694a49a7-0eed-41b5-b595-255e0af53ddb · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Amz driverless: The full autonomous racing system.Journal of Field Robotics, 2020
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b18bc99e-e26c-4b4f-afcb-bc9e5ec90bd2 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Bigger, regularized, optimistic: scaling for compute and sample efficient continuous control.Advances in neural information processing systems, 2024.(Cited on page 6)
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80659e87-a4d6-4f0c-8f44-c83961372503 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6637cae-b873-4075-88ed-f1a08803f236 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Daniel Freeman, Erik Frey, Anton Raichuk, Sertan Girgin, Igor Mordatch, and Olivier Bachem
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7321c3d-81bd-4178-9007-74804c64cf8d · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Unresolved cited work
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9afc4bba-b672-4573-97ea-b3b0ba22dc8a · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Parallel𝑞- learning: Scaling off-policy reinforcement learning under massively parallel simulation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9db694ae-f907-40ab-903c-7895f5901824 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6999d73e-3b90-4573-9448-2d9827b80b98 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5446a11e-446e-4b75-a361-2cd84843bc7c · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 339d6e9c-c26e-4d28-a6e1-2bda3738025a · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots These issues are easy to overlook but they change learning dynamics and final performance
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe94c184-2dd8-4ba4-ae7f-302b94df8581 · outbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots Unresolved cited work
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50d12e37-728d-474c-9076-3b96a36dfde5 · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? What Matters for Simulation to Online Reinforcement Learning on Real Robots
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4164f62f-e63c-4b79-acf8-92dde9862f29 · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? What Matters for Simulation to Online Reinforcement Learning on Real Robots
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f4131a2-19b8-4be1-b740-ba5a28bc8d2f · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? What Matters for Simulation to Online Reinforcement Learning on Real Robots
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c5dd250e-99a1-4000-9ec4-69b9c9d8cff6 · inbound
Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion What Matters for Simulation to Online Reinforcement Learning on Real Robots
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a8e4ee06-e705-480f-8093-511a101ddd32 · inbound
Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors What Matters for Simulation to Online Reinforcement Learning on Real Robots
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.