Pith. sign in

Paper Citation Record · LEDGER

AlgaeDICE: Policy Gradient from Arbitrary Experience

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:1912.02074.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1912.02074 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:33:37.674295Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:33:54.438896Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dcb8eeaf-0b4c-4764-8eac-50538e18f1bf · inbound

D4RL: Datasets for Deep Data-Driven Reinforcement Learning cites this paper.

D4RL: Datasets for Deep Data-Driven Reinforcement Learning AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:19:17.364644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T23:19:17.322890Z digest=sha256:8261c0b48a552256a47ffa1dd0005aef0c5879d40ad63ec8f4c2134cbaf5f4ff

Observation d2a39515-e24d-4a8e-95ce-8949f9310bc1 · inbound

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems cites this paper.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 188

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:21.191873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:db649b97e954b947652417d8b1d2a1d9c2f39a6c80e5940696ab06db56e2c08e

Observation 32019995-bd75-40ec-881c-00b64ac69c3d · inbound

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training cites this paper.

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:42:52.757965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T04:42:52.627166Z digest=sha256:2098e90e99f7aeaf9c6e9724e238a795231a8cba83f0c022cdfe5a5c379e776d

Observation c9c4ecd0-d1ff-49e4-a265-cf6149f78de6 · inbound

TRAM: Test-Time Risk Adaptation with Mixture of Agents cites this paper.

TRAM: Test-Time Risk Adaptation with Mixture of Agents AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:15:49.930546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T22:15:03.665956Z digest=sha256:aa0e88b41ee2ad452907580f83b65dd4c0767ae2f9eeecb2b8b8613fe1b943ca

Observation 4dda242e-b6ad-483e-aaad-dc33477ab5ea · inbound

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets cites this paper.

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:54:21.646502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T20:52:08.062450Z digest=sha256:d29f609124e10397fadc89cd49a06fe17fdd8631d8eeae8cf621dc39f60698a9

Observation 2ac92e50-3ed8-47c2-9e22-a21e3f1f4c93 · inbound

Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting cites this paper.

Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T19:08:19.125336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:05:34.417424Z digest=sha256:646b11cad4d69454911cb049b8e580db1441c80f2390590e1ee0f7e0365822a8

Observation 4bc2a7c4-9efc-49f5-bb6b-2d652d0c73b7 · inbound

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration cites this paper.

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T19:58:23.626145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:53:23.616808Z digest=sha256:979f41aef1559123769d3a2e1eb52abf131d9398536212969117ce381b577c3d

Observation 0514f7e0-b419-4bde-9229-2b6153f84a2d · inbound

Fitted Occupancy-Ratio Evaluation without Bellman Completeness cites this paper.

Fitted Occupancy-Ratio Evaluation without Bellman Completeness AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-07T14:33:54.440537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-07T14:25:33.921937Z digest=sha256:a61e9fa71c4f48e08fb65e64051f7ca07bb0e4ce263cfa62d9e02140e284ec29

Observation 721e975c-8175-4b72-beb9-9e47c0d1c7e0 · inbound

Fitted Occupancy-Ratio Evaluation without Bellman Completeness cites this paper.

Fitted Occupancy-Ratio Evaluation without Bellman Completeness AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T08:33:37.674295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:33:37.674295Z digest=sha256:bc34ca71b7133d84bd911e3a1415c1570fbdcb865240fef78395b16d7f992ab9

Observation 45db64a5-2846-44ee-b5bf-ed1199eeea35 · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 208

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:17.002355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:17.002355Z digest=sha256:a8930efa31c7e1c9e4ab99b81a1fb454f209f7f291ade3225954af4ed6fc86d0

Observation b99b10e1-b87a-4a00-ac19-a794a8110f1e · inbound

Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation cites this paper.

Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation AlgaeDICE: Policy Gradient from Arbitrary Experience

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T13:15:13.026138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:15:13.026138Z digest=sha256:ee969cc6c4cac6033591e7a34c51e3522819edae5453d226ea865ba30eb9e500