Pith. sign in

Paper Citation Record · LEDGER

On Gaussian approximation for entropy-regularized Q-learning with function approximation

As of 4 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2605.17678.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.17678 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-19T22:11:31.021067Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact6
  • verified fuzzy31
  • unresolved2
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40dc8837-f206-4680-b1e1-df640f0feb55 · outbound

This paper cites Residual algorithms: Reinforcement learning with function approximation.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Residual algorithms: Reinforcement learning with function approximation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.813204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:0a886b5671889c83f9633b6c37ee860f2c7ae4dcf33dbc8def816d8bc8416069

Observation e8fd7740-ff10-4615-8216-4925873df3e2 · outbound

This paper cites The reverse isoperimetric problem for gaussian measure.Discrete & Computational Geometry, 10(4):411–420.

On Gaussian approximation for entropy-regularized Q-learning with function approximation The reverse isoperimetric problem for gaussian measure.Discrete & Computational Geometry, 10(4):411–420

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.804765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:2d21ce30ed474df34d46779dd57eabe7b033f7c9d91c6282024d3177db40b055

Observation f7ca4278-657f-4e1b-bae1-41d2c9ffa5bf · outbound

This paper cites Bertsekas and John N.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Bertsekas and John N

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.767640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:f0a210e85fa1566c77bbb3c72a7a51a3ff5087ba4adc1340bfea629782509322

Observation b793ca98-8c8d-4389-b8a0-3146719a43bc · outbound

This paper cites Gaussian approximation for two-timescale linear stochastic approximation.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian approximation for two-timescale linear stochastic approximation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.743242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:76e7658a279e0c724f6a0373724a0caf8f9bfdb33fb7e9edc518fbd579011251

Observation cc8bc86e-9497-4ca0-9079-6039393f0183 · outbound

This paper cites Finite-sample analysis of nonlinear stochastic approximation with applications in reinforcement learning.Automatica, 146:110623.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-sample analysis of nonlinear stochastic approximation with applications in reinforcement learning.Automatica, 146:110623

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.746999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:f485f5b42a69abc393447cf1b926571171b6f8fc06cd65c3f8ea97b23518b2e6

Observation 1e6cdc81-af56-4c0f-8772-7b4556bec9c0 · outbound

This paper cites Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors.The Annals of Statistics, 41(6):2786– 2819.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors.The Annals of Statistics, 41(6):2786– 2819

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.757801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:1bdf89ff84f336ea00067e6860931fdb987181cc310c8e0248ea89c708a02aa4

Observation b4adc09d-effc-437e-a05f-37b3b0c40497 · outbound

This paper cites Springer Series in Operations Research and Financial Engineering.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Springer Series in Operations Research and Financial Engineering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.808946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:129963f0efa9e8bba4abdb84b09a32152b208d9abe524edd399eab871707f340

Observation 953102f2-f448-4b5c-a2f8-7e71798a057c · outbound

This paper cites Finite-time high-probability bounds for Polyak–Ruppert averaged iterates of linear stochastic approximation.Mathematics of Operations Research, 50(2):935–964.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-time high-probability bounds for Polyak–Ruppert averaged iterates of linear stochastic approximation.Mathematics of Operations Research, 50(2):935–964

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.845754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:d72807436318f8cb6ead42d3a67946e6516bcf3e2ccfd7f387d224391fda30ee

Observation df473692-34bd-485f-813a-a448034ce0fb · outbound

This paper cites Tight high probability bounds for linear stochastic approximation with fixed stepsize.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Tight high probability bounds for linear stochastic approximation with fixed stepsize

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.739637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:3e4ee296f1b49f7e26224e0bd352b2bdb3c3a12178ae65e5718d4933c44535a7

Observation 6b744a31-b812-4c34-974b-9bc61127c235 · outbound

This paper cites Online bootstrap confidence intervals for the stochastic gradient descent estimator.Journal of Machine Learning Research, 19(78):1–21.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Online bootstrap confidence intervals for the stochastic gradient descent estimator.Journal of Machine Learning Research, 19(78):1–21

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.736485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:f60df264c73f11d13bbd31b18ad44c231804855e4e89ddc77867d66c710c1ec9

Observation 118e68be-7a8c-4b6b-8384-8da4f7e80917 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Reinforcement learning with deep energy-based policies

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.733271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:f81e38c7387117487a3e53050b173493e4bbacccd80591c524850147b8ec984f

Observation 8a78efbe-f0f9-4a9e-9257-2068de33b18c · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.826715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:76873a1c844f5e0388a89e21be0462f51a2f7cc8a4b4f61a80c9f95eb2fda7bb

Observation 060b5668-ee5c-47c2-8114-ba336ff908b2 · outbound

This paper cites Is q-learning provably efficient? Advances in neural information processing systems, 31.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Is q-learning provably efficient? Advances in neural information processing systems, 31

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.831286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:2bbd009dd08f07f55127e10bf2ccf1cfe93eead96b67b3e06270969b38b8f58c

Observation e5583dc9-0ec9-450f-bbf1-94f4521dd244 · outbound

This paper cites Kaledin, E.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Kaledin, E

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.835145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:cccdeb7b27e8db0fc58c46389fedaace5fa5688bcf697871740fa5aaec0e3396

Observation 209d5770-4039-46a5-bd9b-45f0152652ed · outbound

This paper cites Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1):222–236

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.838768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:e9fe4880f9cf4f15dc3972575c9347ae8b10894d0f8955265c611b72a61f8a29

Observation 8fdd6c56-f3c0-42c5-80aa-07b7ecb3d492 · outbound

This paper cites A statistical analysis of polyak-ruppert averaged q-learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation A statistical analysis of polyak-ruppert averaged q-learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.850007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:a9b4dda019a3d70d360ff685c8f523a57a7e6d8cc05d3ba4bb1243150a7f7f44

Observation 32c7a4a3-69ac-4f90-b804-78312499e4dd · outbound

This paper cites Central Limit Theorems for Asynchronous Averaged Q-Learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Central Limit Theorems for Asynchronous Averaged Q-Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-19T22:12:50.608140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:8fa84606d7c625b547b0545d932fc062169bfe3d7b3ecd8a680d60990c626695

Observation 508b2578-7e2a-4c6c-99fe-e509461b06f3 · outbound

This paper cites An analysis of reinforcement learning with function approximation.

On Gaussian approximation for entropy-regularized Q-learning with function approximation An analysis of reinforcement learning with function approximation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.729484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:c1ef695573040efbb110a642ae53278fdf26e2355977a58844e0a4b104a3bbb4

Observation aec18e48-b3ef-4c81-bc0a-4697f0443765 · outbound

This paper cites Human-level control through deep reinforcement learning.nature, 518(7540):529–533.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Human-level control through deep reinforcement learning.nature, 518(7540):529–533

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.858277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:1e679805236eee0bda56507cb41c8a963f7c2df6676efadfa411e6ffec9e1324

Observation 1524fdd9-9188-4f63-9224-23e5be58d80c · outbound

This paper cites Multivariate normal approximation on the wiener space: new bounds in the convex distance.Journal of Theoretical Probability, 35(3):2020–2037.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Multivariate normal approximation on the wiener space: new bounds in the convex distance.Journal of Theoretical Probability, 35(3):2020–2037

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.786266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:bbcaa4b9291a9e6bb9b26fc321300531ed2d2f2f3c3f0d1801da75dda4caff99

Observation 203198b1-8035-4630-a19b-f4f2075abb3f · outbound

This paper cites Osekowski.Sharp Martingale and Semimartingale Inequalities.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Osekowski.Sharp Martingale and Semimartingale Inequalities

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.790140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:adeed023d8ca7ecc2aec4682154b6368f25eff061fb7999bc2d93cd15e838329

Observation 391bdb96-4866-4296-8740-321abfbab4fe · outbound

This paper cites Concentration inequalities for markov chains by marton couplings and spectral methods.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Concentration inequalities for markov chains by marton couplings and spectral methods

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.800334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:4353448426bf251d8ea943c042858d2f55e9debf7ce4c7571d2999d88da9c70f

Observation 7ecdfc75-5ad1-4064-827a-d3a3a54c8725 · outbound

This paper cites Acceleration of stochastic approximation by averaging.SIAM journal on control and optimization, 30(4):838–855.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Acceleration of stochastic approximation by averaging.SIAM journal on control and optimization, 30(4):838–855

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.818797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:dd6987c7187a62fbda6993cf71e50606838e180d93120b8a0e70ad23d9f7ed38

Observation 08116d00-f759-4f14-a3cd-8f3fb45178e3 · outbound

This paper cites Finite-time analysis of asynchronous stochastic approximation and q-learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-time analysis of asynchronous stochastic approximation and q-learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.782504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:aed9ee340f925ad5e798f754d6dda204d42f3c31722f1219b280c43f38c08766

Observation 67ac2044-f6d0-4135-b51c-555c9ef1d1c8 · outbound

This paper cites Gaussian Approximation for Asynchronous Q-learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian Approximation for Asynchronous Q-learning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-19T22:12:50.631664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:88dbbeaa27915aae11f0da9d3fe95e2c85846612b038b6ad849be861f2e517e7

Observation d6e02392-db8d-41cf-a0d2-2aa2e02dc629 · outbound

This paper cites Efficient estimations from a slowly convergent Robbins-Monro process.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Efficient estimations from a slowly convergent Robbins-Monro process

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.778335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:6f0ccdf35ada1990d9773256a7eb9d4182e5f9721cb24092af6a99e98547a22b

Observation 0c601399-bcdc-4801-8343-901af607ff31 · outbound

This paper cites an unresolved cited work.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-19T22:12:51.793438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:0d4c88ad7e68dbb32cb88045a6667db44fa151f62f967ab847366ff65706f743

Observation f04fa30c-eab2-4156-a128-fb3c51c39b69 · outbound

This paper cites Statistical inference for linear stochastic approximation with Markovian noise.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Statistical inference for linear stochastic approximation with Markovian noise

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.771205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:9f1113e3755837e965ca9660a373f1483237ba8e132454cdaba9196a971bc759

Observation 287755d5-24b7-4c62-95ec-aec4668ecaa3 · outbound

This paper cites Berry–Esseen bounds for multivariate nonlinear statistics with applications to M-estimators and stochastic gradient descent algorithms.Bernoulli, 28(3):1548–1576.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Berry–Esseen bounds for multivariate nonlinear statistics with applications to M-estimators and stochastic gradient descent algorithms.Bernoulli, 28(3):1548–1576

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.753827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:a315f7334a068de0301c9c4b38c2d8d118fedc34b27d6b5ad86575898e4a96ee

Observation 71f9fac5-4741-4776-bb68-87c373122e96 · outbound

This paper cites Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-19T22:12:50.620201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:6f01d64770d5a7fc881ab72b082356c21bb6367e4d6d8bb946359c3592d99855

Observation a3f23be1-9570-484c-9e3a-1392e09de842 · outbound

This paper cites Bootstrap confidence sets under model misspecification.The Annals of Statistics, 43(6):2653 – 2675.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Bootstrap confidence sets under model misspecification.The Annals of Statistics, 43(6):2653 – 2675

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.842199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:a26acfc4f0e49e1007669e0e55a99fb5f308e29afeb4aedb80fdc89acee772f4

Observation 807db67d-64a3-43aa-a5bc-e3e3cb3bfd73 · outbound

This paper cites Sutton and Andrew G.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Sutton and Andrew G

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.822631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:6e0199f0c96ca290b74c72eae0bc9ef0565c25d2a547900532e62ca7b24b20a1

Observation 928aa72a-7791-42a8-a182-61fa991f91a7 · outbound

This paper cites Asynchronous stochastic approximation and q-learning.Machine learning, 16(3):185– 202.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Asynchronous stochastic approximation and q-learning.Machine learning, 16(3):185– 202

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.750525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:dca2e2bc394594041ca7526e1a81159fc70d0bf9a7d6f05adfa3c3dab6ec1574

Observation 0ccccb12-8fc1-4a95-8d3b-35cbb733e050 · outbound

This paper cites Variance-reduced $Q$-learning is minimax optimal.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Variance-reduced $Q$-learning is minimax optimal

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-19T22:12:50.614532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:182a553ae66b2a56c98058386c1106e2a3a42d3163bc23affdd90b93cf004408

Observation e9aa5501-8d7d-48ef-82fe-75326dc04a31 · outbound

This paper cites an unresolved cited work.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-19T22:12:51.774620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:4e1d5edcf75ae7c0c31fecddb7387ac4e33dffcde74a0d8e08f68efc274ab869

Observation f05ad8a7-7a48-4338-98e0-7b9bd8a13c8b · outbound

This paper cites Statistical Inference for Policy Evaluation with Temporal Difference Learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Statistical Inference for Policy Evaluation with Temporal Difference Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:13:09.455495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:454000fab8d53ee21ee062e7a2f229133b5ad3ec708ea36b65e631a54b5a1e0a

Observation db2bb3f6-2fb8-4653-9f95-106f1d1bf2c8 · outbound

This paper cites Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:03:46.373087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:94c88e3335f356f7da8cb2545550c4469055441ac3344b6f0f532f7a96160ca7

Observation 79604292-4892-4c8c-bb2c-b000f8852e1f · outbound

This paper cites A statistical online inference approach in averaged stochastic approxi- mation.Advances in Neural Information Processing Systems, 35:8998–9009.

On Gaussian approximation for entropy-regularized Q-learning with function approximation A statistical online inference approach in averaged stochastic approxi- mation.Advances in Neural Information Processing Systems, 35:8998–9009

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.763463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:67f7556ce19ee6ebb9e3fc13836e1c3940530b147e9c634f99ce8ced31d2d5bc

Observation fe4d978b-8f79-439b-b12e-1b671735b0c2 · outbound

This paper cites Maximum entropy inverse reinforcement learning.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Maximum entropy inverse reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T22:12:51.796753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:f3c7577ee896e1bd47d00adacbba2f9d72e957dcdb3150cb07acd556b1c30ee6

Observation f12d2792-203e-45b6-9acb-b1fe7603dd29 · outbound

This paper cites Finite-sample analysis for sarsa with linear function approximation.Advances in neural information processing systems, 32.

On Gaussian approximation for entropy-regularized Q-learning with function approximation Finite-sample analysis for sarsa with linear function approximation.Advances in neural information processing systems, 32

Reference 40

Resolution
malformed identifier
raw_fallback, observed 2026-05-19T22:12:51.854415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:11:31.021067Z digest=sha256:fcb8f3639c98efdb3f0e8f56a5f44711089dbe313e3c61d62b6ee49119d16709

Pith citing papers

No inbound Pith citation observations are available.