Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T22:11:31.021067Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2605.17678.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T22:11:31.021067Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 40dc8837-f206-4680-b1e1-df640f0feb55 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Residual algorithms: Reinforcement learning with function approximation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8fd7740-ff10-4615-8216-4925873df3e2 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation The reverse isoperimetric problem for gaussian measure.Discrete & Computational Geometry, 10(4):411–420
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f7ca4278-657f-4e1b-bae1-41d2c9ffa5bf · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Bertsekas and John N
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b793ca98-8c8d-4389-b8a0-3146719a43bc · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian approximation for two-timescale linear stochastic approximation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc8bc86e-9497-4ca0-9079-6039393f0183 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-sample analysis of nonlinear stochastic approximation with applications in reinforcement learning.Automatica, 146:110623
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1e6cdc81-af56-4c0f-8772-7b4556bec9c0 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors.The Annals of Statistics, 41(6):2786– 2819
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b4adc09d-effc-437e-a05f-37b3b0c40497 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Springer Series in Operations Research and Financial Engineering
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 953102f2-f448-4b5c-a2f8-7e71798a057c · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-time high-probability bounds for Polyak–Ruppert averaged iterates of linear stochastic approximation.Mathematics of Operations Research, 50(2):935–964
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df473692-34bd-485f-813a-a448034ce0fb · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Tight high probability bounds for linear stochastic approximation with fixed stepsize
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b744a31-b812-4c34-974b-9bc61127c235 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Online bootstrap confidence intervals for the stochastic gradient descent estimator.Journal of Machine Learning Research, 19(78):1–21
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 118e68be-7a8c-4b6b-8384-8da4f7e80917 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Reinforcement learning with deep energy-based policies
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a78efbe-f0f9-4a9e-9257-2068de33b18c · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 060b5668-ee5c-47c2-8114-ba336ff908b2 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Is q-learning provably efficient? Advances in neural information processing systems, 31
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5583dc9-0ec9-450f-bbf1-94f4521dd244 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Kaledin, E
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 209d5770-4039-46a5-bd9b-45f0152652ed · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8fdd6c56-f3c0-42c5-80aa-07b7ecb3d492 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation A statistical analysis of polyak-ruppert averaged q-learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 32c7a4a3-69ac-4f90-b804-78312499e4dd · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Central Limit Theorems for Asynchronous Averaged Q-Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 508b2578-7e2a-4c6c-99fe-e509461b06f3 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation An analysis of reinforcement learning with function approximation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aec18e48-b3ef-4c81-bc0a-4697f0443765 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Human-level control through deep reinforcement learning.nature, 518(7540):529–533
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1524fdd9-9188-4f63-9224-23e5be58d80c · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Multivariate normal approximation on the wiener space: new bounds in the convex distance.Journal of Theoretical Probability, 35(3):2020–2037
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 203198b1-8035-4630-a19b-f4f2075abb3f · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Osekowski.Sharp Martingale and Semimartingale Inequalities
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 391bdb96-4866-4296-8740-321abfbab4fe · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Concentration inequalities for markov chains by marton couplings and spectral methods
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ecdfc75-5ad1-4064-827a-d3a3a54c8725 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Acceleration of stochastic approximation by averaging.SIAM journal on control and optimization, 30(4):838–855
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 08116d00-f759-4f14-a3cd-8f3fb45178e3 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-time analysis of asynchronous stochastic approximation and q-learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67ac2044-f6d0-4135-b51c-555c9ef1d1c8 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian Approximation for Asynchronous Q-learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d6e02392-db8d-41cf-a0d2-2aa2e02dc629 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Efficient estimations from a slowly convergent Robbins-Monro process
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c601399-bcdc-4801-8343-901af607ff31 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f04fa30c-eab2-4156-a128-fb3c51c39b69 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Statistical inference for linear stochastic approximation with Markovian noise
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 287755d5-24b7-4c62-95ec-aec4668ecaa3 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Berry–Esseen bounds for multivariate nonlinear statistics with applications to M-estimators and stochastic gradient descent algorithms.Bernoulli, 28(3):1548–1576
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 71f9fac5-4741-4776-bb68-87c373122e96 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3f23be1-9570-484c-9e3a-1392e09de842 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Bootstrap confidence sets under model misspecification.The Annals of Statistics, 43(6):2653 – 2675
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 807db67d-64a3-43aa-a5bc-e3e3cb3bfd73 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Sutton and Andrew G
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 928aa72a-7791-42a8-a182-61fa991f91a7 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Asynchronous stochastic approximation and q-learning.Machine learning, 16(3):185– 202
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ccccb12-8fc1-4a95-8d3b-35cbb733e050 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Variance-reduced $Q$-learning is minimax optimal
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9aa5501-8d7d-48ef-82fe-75326dc04a31 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f05ad8a7-7a48-4338-98e0-7b9bd8a13c8b · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Statistical Inference for Policy Evaluation with Temporal Difference Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation db2bb3f6-2fb8-4653-9f95-106f1d1bf2c8 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 79604292-4892-4c8c-bb2c-b000f8852e1f · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation A statistical online inference approach in averaged stochastic approxi- mation.Advances in Neural Information Processing Systems, 35:8998–9009
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fe4d978b-8f79-439b-b12e-1b671735b0c2 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Maximum entropy inverse reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f12d2792-203e-45b6-9acb-b1fe7603dd29 · outbound
On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-sample analysis for sarsa with linear function approximation.Advances in neural information processing systems, 32
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
No inbound Pith citation observations are available.