Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions

As of 15 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 3 inbound Pith citation observations for arXiv:2601.15353.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.15353 v2

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T09:09:53.522461Z

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T07:35:56.679877Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T23:49:15.006610Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 38586ef8-4bfb-4e30-a06e-20bc3e6cd2ff · outbound

This paper cites Offline Meta Learning of Exploration.

Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions Offline Meta Learning of Exploration

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T09:09:53.416330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:09:53.416330Z digest=sha256:ba2412424f4a8ed8eaf7e9d8120991bb983e572ccc858e458d67c4c0dfc734ec

Observation 21904c4c-9910-44da-bd85-a2836baf8866 · outbound

This paper cites A Review of Off-Policy Evaluation in Reinforcement Learning.

Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions A Review of Off-Policy Evaluation in Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T09:09:53.522461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:09:53.522461Z digest=sha256:b722f45a5655d8417b44932994b6006c58f290a46dafa49608241eafd1292e07

Pith citing papers

Observation 2c17b700-8c59-459f-b0ac-73ab4422e4b7 · inbound

Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning cites this paper.

Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-14T02:20:22.498368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T07:00:32.206081Z digest=sha256:7cdfd7c4a49f0f0b89cb64429e1916ab0ef1679c0f3b928dc1500474f23ea816

Observation 6d682d5f-2242-4a5e-9255-9b1b1406014f · inbound

Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning cites this paper.

Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-14T02:20:22.498368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T23:47:53.282259Z digest=sha256:7f8f459a1a9acdf0a10089fb741bcd8f09c8296db1ff34ce71c6bf8f5bc087e4

Observation 348f570c-20ab-4a61-817b-31de6590a574 · inbound

A Diffusion-Model Subpopulation Digital Twin for Mobile Health Deployment: A Case Study on the HeartSteps Intervention cites this paper.

A Diffusion-Model Subpopulation Digital Twin for Mobile Health Deployment: A Case Study on the HeartSteps Intervention Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T07:35:56.679877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:35:56.679877Z digest=sha256:2bb19221001ee451ac19f90bf715be2b20aaf111fc2480a88801320355baa50b