Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:1911.06854.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:59:43.198273Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T14:33:54.441902Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation daf5ec59-b9ec-45d2-9b0b-1355b7dab123 · inbound
Off-Policy Evaluation Under Nonignorable Missing Data Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f03b118-fe6d-4584-bc95-413a32e2d38a · inbound
The Three Regimes of Offline-to-Online Reinforcement Learning Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17fb03ed-4e93-4b4e-b7cb-4295feb810d8 · inbound
MedGym:A Unified Continuous-Time Benchmark for Dynamic Medical Treatment Reinforcement Learning Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a77fb6c2-c596-4c43-b9b6-49d309c4b0c2 · inbound
Autoregressive Diffusion World Models for Off-Policy Evaluation of LLM Agents Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 09354f68-833e-42fd-953b-a641e070e521 · inbound
Autoregressive Diffusion World Models for Off-Policy Evaluation of LLM Agents Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 1996
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456e3113-319c-4082-9cf0-80b2ab594428 · inbound
Off-Policy Evaluation for Missingness-Aware Policies in MDPs with Rewards Missing Not at Random Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation acdee659-2dcf-4622-804e-1e19289bd25e · inbound
Fitted Occupancy-Ratio Evaluation without Bellman Completeness Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c5f9debe-2b57-41a2-914f-90e7352b1446 · inbound
Fitted Occupancy-Ratio Evaluation without Bellman Completeness Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.