Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2305.09836.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T11:01:33.933425Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T07:51:16.626435Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation e93d6353-2fa9-4423-ab1c-f0619dffc6d6 · inbound
Value Flows Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d619b10-8368-4f73-a930-90479d29d469 · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ca759100-b47d-4607-8893-591603c2495c · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 990e4c5e-04d6-4a26-940c-ba6d3b720a6e · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a0bd2cc3-7b94-46b1-9f37-a7f338af5164 · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 961bcf7d-6779-45d9-80f6-d5ab4545ef13 · inbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.