Pith. sign in

Paper Citation Record · LEDGER

Gaussian Approximation for Asynchronous Q-learning

As of 4 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2604.07323.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.07323 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T17:11:48.816106Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T22:11:31.021067Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-19T22:12:50.628537Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact14
  • verified fuzzy33
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 17c09e8b-3748-41c0-b2ac-c4d95f901c19 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.883783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:f7f31f9dfb5aeb76d381263cadac570636006cf4d8f214f7402874abf5726502

Observation efe8ea1e-741d-43a6-8ace-af9725d542ed · outbound

This paper cites Anderson, Peter Hall, and D.

Gaussian Approximation for Asynchronous Q-learning Anderson, Peter Hall, and D

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.899053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:f26fb307b290885ec2e82651c45d4885f19db1fa8670300cefd27662ed31ace6

Observation 9fc5c279-16ae-42e9-adb5-bb3d4c942863 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.850252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:ce36eeb9a99a1bb807e1d1d3da6e852aa0373d61563421becd0b5e8e35dd6b64

Observation 06a45f65-b833-41e2-9e0a-4604587f2fe3 · outbound

This paper cites A high dimensional Central Limit Theorem for martingales, with applications to context tree models.

Gaussian Approximation for Asynchronous Q-learning A high dimensional Central Limit Theorem for martingales, with applications to context tree models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T23:02:29.699054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:1df32ff1eb19e2adb6dc4aa479446278db621982b0975978c291f332900ff4d5

Observation f27a4c4f-7c2e-4a61-a8e8-e8f530dde067 · outbound

This paper cites On the dependence of the Berry–Esseen bound on dimension.Journal of Statistical Planning and Inference, 113(2):385–402.

Gaussian Approximation for Asynchronous Q-learning On the dependence of the Berry–Esseen bound on dimension.Journal of Statistical Planning and Inference, 113(2):385–402

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.942188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:f63a8051568c50c4e9073aaaa6072589ee7f32e41e31ea7f6b022c533bcc1958

Observation 07b59856-044a-4407-9677-b4d7d61bedfd · outbound

This paper cites Bertsekas and John N.

Gaussian Approximation for Asynchronous Q-learning Bertsekas and John N

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.921555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:1bb5bc7429728a03dc4214b5b67c930ece758366df5ccffa530c87d01ede1c6e

Observation 2174a691-2286-45f9-be86-002044e564fa · outbound

This paper cites Bhattacharya and R.

Gaussian Approximation for Asynchronous Q-learning Bhattacharya and R

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.914735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:fccbf1349ec9ecaf4b392fca649001bbab0d374e6ae21dbf4ae0f3ab69e1f801

Observation c8f17ddd-5034-4d22-8544-00d4d0520004 · outbound

This paper cites Exact Convergence Rates in Some Martingale Central Limit Theorems.The Annals of Probability, 10(3):672 – 688.

Gaussian Approximation for Asynchronous Q-learning Exact Convergence Rates in Some Martingale Central Limit Theorems.The Annals of Probability, 10(3):672 – 688

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.853698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:5d38bf2cae9b577f8197e8e354212d5b615f3fd0894387fab2d054913ccf374f

Observation a9fdbf16-8686-48ec-84f9-7bebe742bf42 · outbound

This paper cites and ZHANG, Z.-S.

Gaussian Approximation for Asynchronous Q-learning and ZHANG, Z.-S

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:25:59.869875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:5bff082b654abf5490379b9c455f21754d5df38149cf98ef0ca39ff43a22da09

Observation 2aaa03ba-8897-4945-b536-63505eb44f6d · outbound

This paper cites Gaussian approximation for two-timescale linear stochastic approximation.

Gaussian Approximation for Asynchronous Q-learning Gaussian approximation for two-timescale linear stochastic approximation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.891365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:4cc828a295757ea1b1c3a7de1d55e6e4b435edf5cd4e116484037c34c1356ab4

Observation c5bddf65-67f5-4d16-b6df-0f9200439e0f · outbound

This paper cites Lee, Xin T.

Gaussian Approximation for Asynchronous Q-learning Lee, Xin T

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.870505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:cc081b9b14f600d8c8ea0d6e22ccd4c89369ce7c6c843f346c561da4cc2d7759

Observation 72a8ec33-9485-47f9-a6f5-d188f4513004 · outbound

This paper cites Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes.

Gaussian Approximation for Asynchronous Q-learning Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.819891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:064d5c31c9371d7b713f05999323cf4e807b2f7736cc3ca38f9b6d328b335ae6

Observation fcae7216-0bbc-402d-aa0d-7970ebaa2483 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.873543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:866be6c4ce36df906ddc74149719a2cec5ea552fe7724fd9b6e003d68ef53f94

Observation 27a3e51e-bbdd-4927-a4f1-643f37723253 · outbound

This paper cites Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors.The Annals of Statistics, 41(6):2786– 2819.

Gaussian Approximation for Asynchronous Q-learning Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors.The Annals of Statistics, 41(6):2786– 2819

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.895420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:d74ff5da1facc237400874dc506c6c5130e9f028ff465126ce2ec0a7ad801476

Observation f48ccc8a-f047-44f6-8f11-41b400e51b32 · outbound

This paper cites Central limit theorems and bootstrap in high dimensions.The Annals of Probability, 45(4):2309–2352.

Gaussian Approximation for Asynchronous Q-learning Central limit theorems and bootstrap in high dimensions.The Annals of Probability, 45(4):2309–2352

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.880693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:78cc8a1094b21b9350f9e1a7b3aa549fc2a704e682d65043f9de5c142a8e45be

Observation e7971aeb-b4c2-424d-a8cc-5e70df9f0ee2 · outbound

This paper cites Detailed proof of Nazarov's inequality.

Gaussian Approximation for Asynchronous Q-learning Detailed proof of Nazarov's inequality

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.884874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:2a9f2ac9d68bb009fcdfb0342f0016d6259399ffa0aaa694618365456d429f2f

Observation cc2cc32a-19fd-46d1-8f93-ceec902ebe3c · outbound

This paper cites Improved central limit theorem and bootstrap approximations in high dimensions.The Annals of Statistics, 50(5):2562–2586.

Gaussian Approximation for Asynchronous Q-learning Improved central limit theorem and bootstrap approximations in high dimensions.The Annals of Statistics, 50(5):2562–2586

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.867384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:a0a087e549e5855f8a72159cef7004c3c70d3a896933c3c060d856d99d7be0a3

Observation e36df762-04ac-4fc7-a38a-45963a3e7690 · outbound

This paper cites Springer Series in Operations Research and Financial Engineering.

Gaussian Approximation for Asynchronous Q-learning Springer Series in Operations Research and Financial Engineering

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.918303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:2039a6d1f969840bd3d51f9ee18ba5fa59144cc213233e80fc44a923a280bf87

Observation 7311112f-070b-465b-85d7-47e360a6be1b · outbound

This paper cites Learning rates for q-learning.Journal of Machine Learning Research, 5:1–25.

Gaussian Approximation for Asynchronous Q-learning Learning rates for q-learning.Journal of Machine Learning Research, 5:1–25

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.856953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:b7c6f4b86d6af50d27cd14ee0c643e5d8ef5c9b089b7627f1d574b3116419459

Observation a9e0d51c-4f39-4081-9dc9-1d9a44def023 · outbound

This paper cites Exact rates of convergence in some martingale central limit theorems.Journal of Mathematical Analysis and Applications, 469(2):1028–1044.

Gaussian Approximation for Asynchronous Q-learning Exact rates of convergence in some martingale central limit theorems.Journal of Mathematical Analysis and Applications, 469(2):1028–1044

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.935428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:23e58266caee494a1fba11121f45cf7cddd0a3c93e4ba5d7b5c8afa1f20caea6

Observation dc18fa01-dca7-4698-a6e6-fc4a7a2a7122 · outbound

This paper cites High-dimensional central limit theorems by Stein’s method.The Annals of Applied Probability, 31(4):1660 – 1686.

Gaussian Approximation for Asynchronous Q-learning High-dimensional central limit theorems by Stein’s method.The Annals of Applied Probability, 31(4):1660 – 1686

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.925102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:744b6435f246d4f519550e911355955f5d38d62c4265b3db916ca89174a25aba

Observation 0db6baa1-7a9e-4fe1-b0f4-ff81f6943846 · outbound

This paper cites Online bootstrap confidence intervals for the stochastic gradient descent estimator.Journal of Machine Learning Research, 19(78):1–21.

Gaussian Approximation for Asynchronous Q-learning Online bootstrap confidence intervals for the stochastic gradient descent estimator.Journal of Machine Learning Research, 19(78):1–21

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.863805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:8ef60d7579882488a513ac0c1efd4ba29ca7a87e193cc378d5b57e379b27c25b

Observation 34191e91-74e2-4349-90cb-9360d19a440b · outbound

This paper cites Regularity of solutions of the Stein equation and rates in the multivariate central limit theorem.

Gaussian Approximation for Asynchronous Q-learning Regularity of solutions of the Stein equation and rates in the multivariate central limit theorem

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.923408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:00dbed2c306d76e1811a994a41754b8bf8812a34b3c0dcb80006cdeb1755d9ec

Observation ffc5f95e-91e7-475e-a207-56f8c07a08ed · outbound

This paper cites On the rate of convergence in the multivariate CLT.The Annals of Probability, 19(2):724–739, April 1991.

Gaussian Approximation for Asynchronous Q-learning On the rate of convergence in the multivariate CLT.The Annals of Probability, 19(2):724–739, April 1991

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.887791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:b1fe5fa218288ae17f764726bcaa4e5842e12bf3e3690d343b128c33920061f9

Observation b3e73703-63a4-4355-b1ff-935134e12368 · outbound

This paper cites Efficiently solving mdps with stochastic mirror descent.

Gaussian Approximation for Asynchronous Q-learning Efficiently solving mdps with stochastic mirror descent

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.847171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:7bcf0f3905c157107d829fa9ad7a78a7d9d234421a1da00650344d77b435a2ad

Observation e9081292-1292-497b-93fb-dcd0daebc499 · outbound

This paper cites A Berry–Esseen bound for vector-valued martingales.Statistics & Probability Letters, 186:109448.

Gaussian Approximation for Asynchronous Q-learning A Berry–Esseen bound for vector-valued martingales.Statistics & Probability Letters, 186:109448

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.843899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:198621a98069359d5bb3331c2c785821275f408df7a2355e0187ed5ddd01d2a3

Observation 206d7ba6-c205-48c4-9e9d-618f05477532 · outbound

This paper cites Nonasymptotic clt and error bounds for two-time-scale stochastic approximation.arXiv preprint arXiv:2502.09884.

Gaussian Approximation for Asynchronous Q-learning Nonasymptotic clt and error bounds for two-time-scale stochastic approximation.arXiv preprint arXiv:2502.09884

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.844874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:a0b71dfdf284171933ee63429467efdaa86313374e57806868f0f8c19167fc85

Observation dc46c60d-a818-43cc-8323-20572b31ec99 · outbound

This paper cites High-dimensional CLT for Sums of Non-degenerate Random Vectors: $n^{-1/2}$-rate.

Gaussian Approximation for Asynchronous Q-learning High-dimensional CLT for Sums of Non-degenerate Random Vectors: $n^{-1/2}$-rate

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.893122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:c9e06af2869a2cc56560feca504f94937c796c4ec99a68d603a729f92c7864a1

Observation d00910a6-1cc3-4709-8ee1-f0a50e193307 · outbound

This paper cites Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236.

Gaussian Approximation for Asynchronous Q-learning Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.860373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:d87989b8c8bfc2ad79fdc8633e5efb5ebf2ffd186790ad70e69742e57ed7c8dc

Observation af963faf-d7a2-4f0c-b76e-2b2471a5953b · outbound

This paper cites Breaking the sample size barrier in model-based reinforcement learning with a generative model.

Gaussian Approximation for Asynchronous Q-learning Breaking the sample size barrier in model-based reinforcement learning with a generative model

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.903243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:1f4a4220005206c7a0b307705120d6b47c8c549bfe87988c9eb714c89485a740

Observation 1dd25054-6c7f-4863-9902-cc358a299641 · outbound

This paper cites Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236.

Gaussian Approximation for Asynchronous Q-learning Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.911573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:6a95801a26ca1b5679d416fa5a31a9f07da2c2d46b4f00c425bf970333f412ba

Observation 4cf942dd-7f18-4ff9-ba1b-2a7bc596c937 · outbound

This paper cites Asymptotics of Stochastic Gradient Descent with Dropout Regularization in Linear Models.

Gaussian Approximation for Asynchronous Q-learning Asymptotics of Stochastic Gradient Descent with Dropout Regularization in Linear Models

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.852635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:034698aa11d90b11bf39fcbbdca5de726dcd406879aac4d327882ddb97e68403

Observation 68b69d25-12d7-4540-91fb-08739c4433b6 · outbound

This paper cites A statistical analysis of polyak-ruppert averaged q-learning.

Gaussian Approximation for Asynchronous Q-learning A statistical analysis of polyak-ruppert averaged q-learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.938703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:74c851d0c67c2c4302c8de15c6192eb7f8493020c024aac8e29dc84d71c4b981

Observation 319119e5-b75f-4a34-a2f3-45f1d8e6d5d9 · outbound

This paper cites Central Limit Theorems for Asynchronous Averaged Q-Learning.

Gaussian Approximation for Asynchronous Q-learning Central Limit Theorems for Asynchronous Averaged Q-Learning

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:25:59.828468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:e2684e978218bb260531f403616fce555ff3f22b6b20ccb0649caf9f0f802bb2

Observation d353ee5a-e83b-43d6-87ac-192ab437c281 · outbound

This paper cites Concentration inequalities for sums of markov-dependent random matrices.Information and Inference: A Journal of the IMA, 13(4):iaae032.

Gaussian Approximation for Asynchronous Q-learning Concentration inequalities for sums of markov-dependent random matrices.Information and Inference: A Journal of the IMA, 13(4):iaae032

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.837435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:2806d55a61d42c67a048ba86b8ff44c2c56fafa2ccf9eec1317f69910be7ff60

Observation 18d51853-e684-4487-b01b-29f26b6f041f · outbound

This paper cites New stochastic approximation type procedures.Automat.

Gaussian Approximation for Asynchronous Q-learning New stochastic approximation type procedures.Automat

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.817522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:da48f291305bc0045731e70fe02d5cd3626a71126e3aac9b4d9f59fc10864bef

Observation 23324682-fef6-4752-9ddf-5b1b38cd4e22 · outbound

This paper cites Acceleration of stochastic approximation by averaging.SIAM journal on control and optimization, 30(4):838–855.

Gaussian Approximation for Asynchronous Q-learning Acceleration of stochastic approximation by averaging.SIAM journal on control and optimization, 30(4):838–855

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.877292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:9c99786991b3eb45c5ee5183c1ccd6d9b6293c9baaba9a4ddfd1f9b7f945b6d7

Observation 2fab149e-a821-4c51-88b6-a30c5bcd5753 · outbound

This paper cites Puterman.Markov Decision Processes: Discrete Stochastic Dynamic Programming.

Gaussian Approximation for Asynchronous Q-learning Puterman.Markov Decision Processes: Discrete Stochastic Dynamic Programming

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.827624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:ea06312a3c34ce3ea47e79b363c6088cf4fd6e29c9684df72452a7c2140e258e

Observation 92e1848a-d317-4b0a-9530-e3648ff551ac · outbound

This paper cites Finite-time analysis of asynchronous stochastic approximation and q-learning.

Gaussian Approximation for Asynchronous Q-learning Finite-time analysis of asynchronous stochastic approximation and q-learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.820899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:12f94b5a313b612357be1ed97cb9cbc7c68763f918c02a9a4705da944426d0d9

Observation 72acc8df-8150-475e-94b4-825c14445af3 · outbound

This paper cites Multivariate normal approximation with Stein’s method of ex- changeable pairs under a general linearity condition.The Annals of Probability, 37(6):2150 – 2173.

Gaussian Approximation for Asynchronous Q-learning Multivariate normal approximation with Stein’s method of ex- changeable pairs under a general linearity condition.The Annals of Probability, 37(6):2150 – 2173

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.810944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:a91970ed0aa2b55a652159c4a7bc8aaaaf5122949891fe031e85128cefb73c19

Observation 3cb1514c-1381-47dd-a479-190ec1df18e8 · outbound

This paper cites On quantitative bounds in the mean martingale central limit theorem.Statistics & Probability Letters, 138:171–176.

Gaussian Approximation for Asynchronous Q-learning On quantitative bounds in the mean martingale central limit theorem.Statistics & Probability Letters, 138:171–176

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.807530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:f9dfc2cbb1e75c54c54e84b3e324251cbbdc21e2d5b845557e673f114c5c0538

Observation 6149379c-880e-47c1-b7f9-126935c3de6d · outbound

This paper cites Efficient estimations from a slowly convergent Robbins-Monro process.

Gaussian Approximation for Asynchronous Q-learning Efficient estimations from a slowly convergent Robbins-Monro process

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.814172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:37aa53d3c50737e86317cdd56b01c3bb680f9c65591c4a53353f04059eb97cb0

Observation e886c817-3bba-4a6f-b4ce-56ede99313a6 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.824169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:8e136b09cea132b2972f24826eab91ffe8d80d8c3bd3e20e1a9cd170ff911ef0

Observation 5b64ad4c-5317-43c5-ab7e-eeb96956a16b · outbound

This paper cites Statistical inference for Linear Stochastic Approximation with Markovian Noise.

Gaussian Approximation for Asynchronous Q-learning Statistical inference for Linear Stochastic Approximation with Markovian Noise

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.943910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:e01ecb3f7a9ade868398fb282f1108f251aa24f23f7d22c3cd67da9636ed7871

Observation 2c59474f-08a3-450c-a8ca-caa6e2161aae · outbound

This paper cites Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent.

Gaussian Approximation for Asynchronous Q-learning Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:25:59.959150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:3647f52cff48b295bf412fb2dc2792c9495df04d4321cbcb26b0a3e4d090062a

Observation cef8a62f-9158-4d6d-adfc-c035449027a9 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.931896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:a6530ad6025141bb9e29de042d52d389b28aab4ee3e3b6b419c310f04b51bb09

Observation d5213755-02f7-40a6-9bee-2612eef29b20 · outbound

This paper cites Sutton and Andrew G.

Gaussian Approximation for Asynchronous Q-learning Sutton and Andrew G

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.907798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:c7da4512ec9a6a1b2a4ea8ca9ecc8f81afbdc31e8948796038c6e19e6057dcfc

Observation 0a7d760a-9cf2-497d-8880-cb0b59df6dba · outbound

This paper cites Stochastic approximation with cone-contractive operators: Sharp $\ell_\infty$-bounds for $Q$-learning.

Gaussian Approximation for Asynchronous Q-learning Stochastic approximation with cone-contractive operators: Sharp $\ell_\infty$-bounds for $Q$-learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.858264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:bb3790742dc348259ae0da5c4aed59866fca12b31d2fb1b4999c6605bb1bd1ee

Observation d976f237-ad61-48dd-99a0-9f29ce3c1dce · outbound

This paper cites Variance-reduced $Q$-learning is minimax optimal.

Gaussian Approximation for Asynchronous Q-learning Variance-reduced $Q$-learning is minimax optimal

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.876792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:ecbf1cdc60a597d720475df994e691e9584603d84a8701f502c51960c925198d

Observation 754525f4-4a2f-4227-a21d-f8bf142b4f5e · outbound

This paper cites Randomized linear programming solves the markov decision problem in nearly linear (sometimes sublinear) time.Mathematics of Operations Research, 45(2):517–546.

Gaussian Approximation for Asynchronous Q-learning Randomized linear programming solves the markov decision problem in nearly linear (sometimes sublinear) time.Mathematics of Operations Research, 45(2):517–546

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.928609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:10ca723927c8db9b4124011b206c8ac458deb5c98946e0d5fd864f1ce216802e

Observation b53fc583-0aa2-4f6e-abc6-cac615f71a09 · outbound

This paper cites Online Covariance Matrix Estimation in Stochastic Gradient Descent.Journal of the American Statistical Association, 118(541):393–404.

Gaussian Approximation for Asynchronous Q-learning Online Covariance Matrix Estimation in Stochastic Gradient Descent.Journal of the American Statistical Association, 118(541):393–404

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T11:39:32.831150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:068c8682c0ea96ee37ee39927e0dd1f047721554012ef2a3bbc82570ced5a5a8

Observation 11d6687f-2c79-4e2c-9553-fa28fa821a5b · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.834170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:e83131f47793f9b4913ad2c31ab41941b1096b5c4f5000a14e7ef38b20fef264

Observation 361b5503-949d-427c-8362-a47d7f929d60 · outbound

This paper cites an unresolved cited work.

Gaussian Approximation for Asynchronous Q-learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-17T11:39:32.840631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:cebeba1b1a24de7f7c739e1bc4cc18fd63df52b63764a718e5864322c9d3cb26

Observation ae2c99bd-b6bf-481e-9378-71930048d62f · outbound

This paper cites Statistical Inference for Policy Evaluation with Temporal Difference Learning.

Gaussian Approximation for Asynchronous Q-learning Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:fe224b545946b7dcc51ba6755f4b835f992bccce85970390e5b0348d10ae355e

Observation 08382349-42f3-4684-951c-2d12d3197992 · outbound

This paper cites Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning.

Gaussian Approximation for Asynchronous Q-learning Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:03:46.373087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:11:48.816106Z digest=sha256:529612980372c00069ab4ee15f1125dc5006e33ac93d363ebc90808cd122fa92

Pith citing papers

Observation 67ac2044-f6d0-4135-b51c-555c9ef1d1c8 · inbound

On Gaussian approximation for entropy-regularized Q-learning with function approximation cites this paper.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian Approximation for Asynchronous Q-learning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-19T22:12:50.631664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:88dbbeaa27915aae11f0da9d3fe95e2c85846612b038b6ad849be861f2e517e7