Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2305.15669.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:19:29.370866Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T01:33:27.193926Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a8f851a0-5087-40e1-879a-1118306b0361 · inbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a521ed32-5016-4189-875c-5695f1d06c55 · inbound
Online Pre-Training for Offline-to-Online Reinforcement Learning PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 227a11f2-c231-4570-be1f-2d4df3c6b0d4 · inbound
The Three Regimes of Offline-to-Online Reinforcement Learning PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70215d69-c80b-4ef0-b858-61618f8bed27 · inbound
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.