Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:00:10.690877Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2510.01460.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:00:10.690877Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:15:07.735336Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T19:30:07.874682Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8ca2e4c2-6779-4f38-9a1b-28e2cd11ad15 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4474f99-99c2-45de-a42b-e308f36073c0 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Efficient online reinforcement learning with offline data
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a974afb-4e9d-4839-b5a1-c3835a82ee5a · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Data quality in imitation learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64f74df9-265f-46b9-b974-f7982ede9b8b · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Magnetic control of tokamak plasmas through deep reinforcement learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88633603-02c1-4fb2-8bf1-05516eefb654 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Loss of plasticity in deep continual learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5456f005-3387-4bba-9222-69cef5bd08d8 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44f3e1ee-45b7-44d9-bb9a-83644819cb5a · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Addressing function approximation error in actor-critic methods
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d6f22f0-1bc3-4bda-8214-587658d9c3cc · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daeae038-94a1-45bd-99c4-7e163bc04997 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Bayesian design principles for offline-to-online reinforcement learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c19f7a39-8f59-442e-b3e0-843f87283869 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Overcoming catastrophic forgetting in neural networks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79bd1747-8320-4789-9715-1e406d61c846 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcbcfc6b-5b27-45be-8956-82845384997d · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fa3a69c-1c3a-49d6-abe8-1b2a4bd32e81 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 227a11f2-c231-4570-be1f-2d4df3c6b0d4 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 313ddf1d-bd00-4731-b934-2c4038123e26 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Energy-guided diffusion sampling for offline-to-online reinforcement learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f94a57a2-8bdb-42a5-9f85-44c5b4c064e3 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be610524-8dd9-4aac-8c6c-12f6ca48f3e0 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning The stability-plasticity dilemma: Investigating the continuum from catastrophic forgetting to age-limited learning effects
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39ffdfd2-a87e-4575-be96-2b4cc70eb862 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Human-level control through deep reinforcement learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e7b98e4-c3d8-4ef8-ae37-279828667c32 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31d5fcf0-dece-4f4c-a9e9-734ecd514799 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Cal-ql: Calibrated offline rl pre-training for efficient online fine-tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a7bae6b-e680-4c8d-833e-94670839269d · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning The primacy bias in deep reinforcement learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400e34b3-4842-4edf-9fd4-020fae7fdf4c · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning An algorithmic perspective on imitation learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77b01d49-ec9d-4562-8b06-0dd9590a8577 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Experience replay for continual learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d49d5e7-b6d6-4fac-a37b-93a914780dae · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Progressive Neural Networks
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f49e066d-df2a-47e3-8eaa-e10d4939dc0c · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Learning from demonstration
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8baebfc2-c194-436e-b632-c7aab6fa00e4 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Mastering the game of go without human knowledge
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 746071d9-dffe-433a-9a4e-3d668e546014 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning The dormant neuron phenomenon in deep reinforcement learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61191bb3-ed59-4c20-bad3-3192f7dd2288 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fce4ee65-e198-4c01-94b6-eaafdcf147e3 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Feedback in Imitation Learning: The Three Regimes of Covariate Shift
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 506f1782-cc61-4490-ae3c-baa7ff54259b · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Revisiting the minimalist approach to offline reinforcement learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78036c47-8bf7-4ad0-9557-ded2448a01dc · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Jump-start reinforcement learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f03b118-fe6d-4584-bc95-413a32e2d38a · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcb8cee3-b4ea-4d9f-a664-898332447b5b · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Fine-tuning reinforcement learning models is secretly a forgetting mitigation problem
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f83bd745-f59f-48bc-8cb9-c5b98b025e48 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Policy expansion for bridging offline-to-online reinforcement learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 960fc39a-6c78-45a9-8cd2-88d8b797036e · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc4c0827-f1c4-4e30-ab56-44fee2f15ec0 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning @esa (Ref
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0c5250f-5091-4b8a-8cb0-da5f49b64545 · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d42a8413-8f1b-4bd8-b03b-71a19c49ac2d · outbound
The Three Regimes of Offline-to-Online Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa959f4-104c-47cf-becb-304ceca00e76 · inbound
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization The Three Regimes of Offline-to-Online Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c25873e5-3c82-4fe2-b2c1-a573384893d3 · inbound
Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors The Three Regimes of Offline-to-Online Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.