Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:1912.02074.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T08:33:37.674295Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T14:33:54.438896Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation dcb8eeaf-0b4c-4764-8eac-50538e18f1bf · inbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d2a39515-e24d-4a8e-95ce-8949f9310bc1 · inbound
Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 188
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32019995-bd75-40ec-881c-00b64ac69c3d · inbound
VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c9c4ecd0-d1ff-49e4-a265-cf6149f78de6 · inbound
TRAM: Test-Time Risk Adaptation with Mixture of Agents AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4dda242e-b6ad-483e-aaad-dc33477ab5ea · inbound
Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ac92e50-3ed8-47c2-9e22-a21e3f1f4c93 · inbound
Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4bc2a7c4-9efc-49f5-bb6b-2d652d0c73b7 · inbound
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0514f7e0-b419-4bde-9229-2b6153f84a2d · inbound
Fitted Occupancy-Ratio Evaluation without Bellman Completeness AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 721e975c-8175-4b72-beb9-9e47c0d1c7e0 · inbound
Fitted Occupancy-Ratio Evaluation without Bellman Completeness AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45db64a5-2846-44ee-b5bf-ed1199eeea35 · inbound
Reinforcement Learning: From Algorithms To Foundation Models AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 208
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b99b10e1-b87a-4a00-ac19-a794a8110f1e · inbound
Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.