Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:27:02.708438Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 3 inbound Pith citation observations for arXiv:2507.08387.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:27:02.708438Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T01:05:48.492775Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:59:57.315756Z
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 70448eb6-02dc-4ca0-8d32-b2962ff27521 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8add5fa-25f2-47b3-a186-b4b2f4dd8c22 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 139326c2-25fc-4f29-9fc1-fc61249429a7 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Behavior Regularized Offline Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7981c2a-9b58-4e68-ae9f-df17a4c05f3b · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cef0550a-dc3a-48f4-bc45-579eaf9092e5 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 767f32c3-af73-45db-9a21-063a76a9d3bb · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e179b792-806c-4c03-b1dc-e6abe6b28564 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Off-policy deep reinforcement learning without exploration
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a521ed32-5016-4189-875c-5695f1d06c55 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e0cd23f-e002-4ffd-83a0-31b9d45c059b · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Layer Normalization
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a44186e-6c90-44c6-af47-8afd10ca96dc · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbc7a964-c092-41de-8e68-022358a47638 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning Bayesian Design Principles for Offline-to-Online Reinforcement Learning
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09342699-cb04-4d35-88c8-39d28b86cf15 · outbound
Online Pre-Training for Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c21f1114-17c6-47b2-9d84-4946eeee4052 · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Online Pre-Training for Offline-to-Online Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e7b10ab-4a70-45b0-92d6-eedf3ac1f6cc · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Online Pre-Training for Offline-to-Online Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2bddfb0-1199-487d-93f3-8229f9b8363b · inbound
Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources Online Pre-Training for Offline-to-Online Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.