Pith. sign in

Paper Citation Record · LEDGER

Statistical Inference for Policy Evaluation with Temporal Difference Learning

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2410.16106.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.16106 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:29:39.522444Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T03:32:28.324798Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0c02fb04-f828-41c1-93de-eac791acd6b2 · inbound

Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent cites this paper.

Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T03:29:27.821389Z digest=sha256:4bcf774976968407d5a49b966b3ca6194471a847e4c75e632745a692992ce5f5

Observation 8a321498-56ae-4e11-9534-dde4e516619c · inbound

Statistical inference for Linear Stochastic Approximation with Markovian Noise cites this paper.

Statistical inference for Linear Stochastic Approximation with Markovian Noise Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:39.522444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:39.522444Z digest=sha256:1274e1819d3dc141d644f37d7b49802784b6bc2918db5dc2d769b47ff3242013

Observation 3b2fd979-3cbb-4f7f-9ec6-3eec4d9247e5 · inbound

Statistical and Algorithmic Foundations of Reinforcement Learning cites this paper.

Statistical and Algorithmic Foundations of Reinforcement Learning Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.942908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.942908Z digest=sha256:bfc33507cf8447d083c5b68be0dc1ae2d6d29054c7b14559f4aef51afaac52b8

Observation e9787906-6a60-4791-a668-1dbe51521b39 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:1b5cd63f141c56ff2e3e59cf0832231b069583e72767359a3f31acfe884c46a3

Observation ae2c99bd-b6bf-481e-9378-71930048d62f · inbound

Gaussian Approximation for Asynchronous Q-learning cites this paper.

Gaussian Approximation for Asynchronous Q-learning Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:4b00196e37fdf6fafcd9316501851dd722379c20ed688fc93b6e4d1d87f92ca8

Observation f05ad8a7-7a48-4338-98e0-7b9bd8a13c8b · inbound

On Gaussian approximation for entropy-regularized Q-learning with function approximation cites this paper.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:859ff5cef9b0e99c8911bd7384833e754fc1ee488d596cb4df14ea0ccee0b656

Observation a93a39e6-004c-4902-a534-215ae73c971b · inbound

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation cites this paper.

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T02:10:02.167114Z digest=sha256:3a25e50fc4417d67d4acf4dfa63978655c7c62c306e9eb65b920bd248bc9041e