Pith. sign in

Paper Citation Record · LEDGER

Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2401.13884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.13884 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T10:46:25.335756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:05:45.530664Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7059be44-e7d7-47d7-b37c-ac070274758a · inbound

From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes cites this paper.

From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:35:00.842255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T17:34:49.191496Z digest=sha256:f85238673d58e49b45b641524a1d320c013db4d2fed2ff13ba2493cb4ed5d931

Observation ea9749e5-69e4-4847-9ef0-ed364951065e · inbound

Central Limit Theorems for Asynchronous Averaged Q-Learning cites this paper.

Central Limit Theorems for Asynchronous Averaged Q-Learning Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:56:30.395591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T14:55:49.034505Z digest=sha256:5a740b454b01b03f15041d584be4b3e49bbf61a84e780287248794eafa94d1ab

Observation 97155cf2-5cfc-4627-92be-fd8d4a93244e · inbound

A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies cites this paper.

A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T05:50:57.154573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T05:47:47.782246Z digest=sha256:01d4d8d167da6ba584d3cfad770557820197c3aed61dcccb27a106fba7355fc6

Observation 138a38e0-9a1b-42ef-8472-7956e3581bb7 · inbound

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains cites this paper.

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 76

Resolution
malformed identifier
arxiv_id, observed 2026-05-16T15:23:02.079774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T15:21:02.872523Z digest=sha256:50f9e734b72204a96648aa048d42a434012c790dc669fa859519089d77106f8b

Observation 82258d0c-a47d-4238-99e1-9b2d32e949e7 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:bed78a03a304a978b712e79e0c04a6c086ee1b6646ae398a39e32b8bfbf89713

Observation 93b7c764-836c-4c16-80d8-d05af1def6a2 · inbound

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD cites this paper.

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:01:03.712977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-10T15:13:59.395799Z digest=sha256:bc42574055d09da2afae968f3dd0a33fc42d058118e476645b9959715f83e760

Observation 7d748d63-82fb-49f6-8a9b-afccb39e68aa · inbound

Revisiting the Constant Stepsize Stochastic Approximation with Decision-Dependent Markovian Noise cites this paper.

Revisiting the Constant Stepsize Stochastic Approximation with Decision-Dependent Markovian Noise Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:45:27.427927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-10T13:45:04.730863Z digest=sha256:c7fa264a5b76d3285e1c2251d494fdba9789832cc8c70dbb4590092034b51006

Observation cca55c52-38fc-4b2f-8279-dea5dc13086f · inbound

Elephant random walk with attributed steps and extractions of random sizes cites this paper.

Elephant random walk with attributed steps and extractions of random sizes Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:11:20.120298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T06:09:29.594763Z digest=sha256:36273e324ae87e8527a91750208b9882c5f05856c52f0f13b19c6f4395f06df2

Observation a9ac911b-bb77-45f5-85b8-1191a1b91f4c · inbound

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation cites this paper.

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T02:12:58.456482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T02:10:02.167114Z digest=sha256:ca86b1cf763257a13f5577a3f5552f31db2cbfe2510f24595495d25ca6a2bb74

Observation 429ac6e2-176d-4cb7-80e2-78b514eb8ed3 · inbound

SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates cites this paper.

SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 221

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:05:45.532477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-07-01T01:05:17.842447Z digest=sha256:7a49823194984bc7f861c3f867aeab9b06c4f65ad1c870ff20c7c002e15c3b32