Pith. sign in

Paper Citation Record · LEDGER

Statistical and Algorithmic Foundations of Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 100 of 162 outbound references and 1 inbound Pith citation observation for arXiv:2507.14444.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14444 v1

Coverage vector

measured 100 of 162 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:15:15.850895Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T10:46:25.335756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 162 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved89
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 12693451-7cd5-408e-802e-84f3cd7406ac · outbound

This paper cites CS Dept., UW Seattle, Seattle, WA, USA, Tech.

Statistical and Algorithmic Foundations of Reinforcement Learning CS Dept., UW Seattle, Seattle, WA, USA, Tech

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.568377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.568377Z digest=sha256:b03e01c7464c1ef296547eb0aa5abe83e791a681c02071e78bb1124269160b53

Observation 2fe14251-ff17-4bc4-8c1c-dd0d81d322b4 · outbound

This paper cites Advances in neural information processing systems 33:20095–20107.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 33:20095–20107

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.572431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.572431Z digest=sha256:5f45eaa8a223ef14fcb9177da1b9b7e475188edea13ec4604c5ed731c20c9e6e

Observation 1a1d39a8-6379-45e1-8724-3890a2d9e279 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.575782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.575782Z digest=sha256:d832a6a2a5b00533cc1e1c1da690837f277776df1d617eb48a22d3ab83938c0b

Observation b58f16cd-094b-42d3-8857-eaf9469b8cba · outbound

This paper cites Journal of Machine Learning Research 22(98):1–76.

Statistical and Algorithmic Foundations of Reinforcement Learning Journal of Machine Learning Research 22(98):1–76

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.579021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.579021Z digest=sha256:43fd89e632b3b615fbf28b0c86a4e36ec30e790d5b083dad658f76807b8e7744

Observation 218e35f8-f1b0-4007-8213-c796eb9579e0 · outbound

This paper cites Advances in neural information processing systems 19.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 19

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.582596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.582596Z digest=sha256:e45d10729ccf53b61850033da87e1d8f67fff62107984b40289ef36613cf290c

Observation 164cdbcc-c8bf-408d-9b11-c58544263029 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.585913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.585913Z digest=sha256:a2ec0d4c171124a89cf77962c4c9b9c89bdec2b5c6f3b29780fa571c16e0fb7d

Observation 3c93fe29-13c3-4d98-9759-9ca2f9be3364 · outbound

This paper cites International Conference on Machine Learning , 263–272 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 263–272 (PMLR)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.589749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.589749Z digest=sha256:3b42887521f80551312814fc1018c7428787896ae3a3d962cc3df5eec882505e

Observation ac69a5cc-b15f-47a4-a5f4-68c53b0d3277 · outbound

This paper cites Systems & control letters 61(12):1203–1208.

Statistical and Algorithmic Foundations of Reinforcement Learning Systems & control letters 61(12):1203–1208

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.592467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.592467Z digest=sha256:e8ae9a56b82b7f5914fef909c2daa2dbe3a2790edd19bb043314115423b01ce0

Observation 45717872-6b3d-4d15-85ee-92d39de1db7d · outbound

This paper cites Proceedings of the National Academy of Sciences of the United States of America 38(8):716.

Statistical and Algorithmic Foundations of Reinforcement Learning Proceedings of the National Academy of Sciences of the United States of America 38(8):716

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.595603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.595603Z digest=sha256:39a4c9be3d31699fe0c0f4508044c7f4ceca9846d27a7c7b2903359e0b243785

Observation c10e113f-0523-4bf7-a256-8d9be687dd17 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.598900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.598900Z digest=sha256:2c25edfdca6b18044f0696a4c01fed7196afeae1a0c8ad563851725b33becf1c

Observation c1ccb9a8-96e6-4b7f-8424-9ce10fcd2a67 · outbound

This paper cites Mathemat- ical Programming 167(2):235–292.

Statistical and Algorithmic Foundations of Reinforcement Learning Mathemat- ical Programming 167(2):235–292

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.601779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.601779Z digest=sha256:a8bd2162cc4560113809c9c3bbe14f494ac8aacffc3aeea7d1d971ee25f31b0f

Observation 619c8f36-1be3-42f9-8a02-fa61d17b76c7 · outbound

This paper cites Management Science 65(2):604–618.

Statistical and Algorithmic Foundations of Reinforcement Learning Management Science 65(2):604–618

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.604774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.604774Z digest=sha256:541a2c78cc2aa0a0203a63c614afa69ce4515bd2214054e32137edfeb1dbb336

Observation 5573b5ee-9447-4ea7-adf7-bd74db139714 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , 2386– 2394 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Artificial Intelligence and Statistics , 2386– 2394 (PMLR)

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.607513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.607513Z digest=sha256:2caa3c0195d72f7692e3a4dde250ae39f9c5656b4757ee341ff305970a3156d1

Observation ce827f2d-70fa-45a4-8dfc-88488d2f0656 · outbound

This paper cites Operations Research 72(5):1906–1927.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 72(5):1906–1927

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.610598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.610598Z digest=sha256:9d0a7bcbac134b73edee7e453c9d5c624cec4a3757e61d4e3f9f1dbbdf28352b

Observation 38973305-f761-42f6-a3af-6a7f523857a4 · outbound

This paper cites Operations Research 69(3):950–973.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 69(3):950–973

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.613430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.613430Z digest=sha256:d4c02badd9669953423c9d09d6f2fd1a287cbc92a19a4405e21d6048a630ad44

Observation 0b98b1da-9ce1-45e8-9cd0-0fac68ea4ee0 · outbound

This paper cites Mathematics of Operations Research 44(2):565–600.

Statistical and Algorithmic Foundations of Reinforcement Learning Mathematics of Operations Research 44(2):565–600

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.616000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.616000Z digest=sha256:326b8293a53e78074d04222d516f5083ba86d927c69a1c7d45bef5cc418f2924

Observation f12e40e0-a906-487f-8a2f-8834cf4e48cb · outbound

This paper cites International Conference on Machine Learning , 1056–1066 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 1056–1066 (PMLR)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.618901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.618901Z digest=sha256:508d0cd81af74b169b3d08d9b90055b2036d78dd87f1c395264bed9c2a560c86

Observation e5f4bb39-e95a-4a5e-92c9-e9163c58fb12 · outbound

This paper cites the method of paired comparisons.

Statistical and Algorithmic Foundations of Reinforcement Learning the method of paired comparisons

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.621807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.621807Z digest=sha256:20f434cfdb9f0b656fcd36de4515de59b298c69e286bbf449d15b636086b7fe4

Observation 5c900be5-fdee-45a8-bf66-a84061a09017 · outbound

This paper cites 2022 IEEE 61st Conference on Decision and Control (CDC) , 2833–2838 (IEEE).

Statistical and Algorithmic Foundations of Reinforcement Learning 2022 IEEE 61st Conference on Decision and Control (CDC) , 2833–2838 (IEEE)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.624549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.624549Z digest=sha256:78ba47ecc3ad4484a80455a0c3bcbfc8b4cf6fafd4541b7493c01f2756268eb1

Observation 4aa3a470-cb49-4e0c-81d6-19c269613c8f · outbound

This paper cites Operations Research 70(4):2563– 2578.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 70(4):2563– 2578

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.627194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.627194Z digest=sha256:b2eb24ab97a5ac5c61ccf8e7adbe088e53136372c7f7d4d170cc084a8d8324c5

Observation 17a364c8-80c2-4674-ba99-80716f2b0381 · outbound

This paper cites The Eleventh International Conference on Learning Representations.

Statistical and Algorithmic Foundations of Reinforcement Learning The Eleventh International Conference on Learning Representations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.630359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.630359Z digest=sha256:814d16a55d2eb2569698e664b5110a80c4fb17c85efcfc21977567be9e34d746

Observation bc4cfaa9-d03e-40ab-9887-1a87081dfdd6 · outbound

This paper cites The Thirteenth International Conference on Learning Representations.

Statistical and Algorithmic Foundations of Reinforcement Learning The Thirteenth International Conference on Learning Representations

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.633188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.633188Z digest=sha256:009de95957ee10bf5ff3f813a76b650375609542471fcb8e2a4c56d65635ee79

Observation 48b82d19-3fb8-4c84-9469-4077e9b9aa8a · outbound

This paper cites Journal of Machine Learning Research 25(4):1–48.

Statistical and Algorithmic Foundations of Reinforcement Learning Journal of Machine Learning Research 25(4):1–48

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.636473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.636473Z digest=sha256:1edced1ae0c06771bd6a3069c1731173b53c63213e0af6e4d5846cff3bff44cb

Observation 77042f5f-0b78-404d-a29c-ea7c816de9fd · outbound

This paper cites Advances in neural information processing systems 30.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 30

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.639126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.639126Z digest=sha256:d21e082c9f266ce428fff93a2db2bb5fc06386763065cb32bbb828dc2d8a8438

Observation 5b3b6c3f-76ee-442d-833c-852d9512156a · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.641900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.641900Z digest=sha256:aaf6ebd7e88558d43fa2ba7280d5f8c9863349e2a95d5faef4ea6ecfe855e5a2

Observation dcb15948-03b4-4caa-861f-8901ab6364a3 · outbound

This paper cites The Annals of Statistics 53(1):426–456.

Statistical and Algorithmic Foundations of Reinforcement Learning The Annals of Statistics 53(1):426–456

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.645160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.645160Z digest=sha256:6675fc9233bb6763c084618e669c4d1aa00ee70d7ee4ec741ade50e996432150

Observation 65d94137-b842-411f-88cb-6cd8608abd07 · outbound

This paper cites International Conference on Machine Learning , 1042–1051 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 1042–1051 (PMLR)

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.647900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.647900Z digest=sha256:1458a3f2163975573f22e5c27ff818b0d818d65cd07de268cfcea3722a497e8f

Observation 7cc32cb6-b0b1-4ecb-baf2-41192b31f171 · outbound

This paper cites Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes.

Statistical and Algorithmic Foundations of Reinforcement Learning Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.650820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.650820Z digest=sha256:f5c416a23fa8543f43a1fe08a679fd574d05633f0562c5bbc40261604c6db681

Observation 067d16b7-c757-4fb4-ad9c-3064f461fe02 · outbound

This paper cites A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants.

Statistical and Algorithmic Foundations of Reinforcement Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.653672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.653672Z digest=sha256:c04f52e423299b0ff3d9011c41860bc855eefe867679dbcbfd69427fbaaca31e

Observation ddfc3f08-603d-42f4-be74-81ef18d9c46b · outbound

This paper cites A Non-Asymptotic Theory of Seminorm Lyapunov Stability: From Deterministic to Stochastic Iterative Algorithms.

Statistical and Algorithmic Foundations of Reinforcement Learning A Non-Asymptotic Theory of Seminorm Lyapunov Stability: From Deterministic to Stochastic Iterative Algorithms

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.656586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.656586Z digest=sha256:a2bc371c61ab903d9697aba8e222db45c89af2df3118b2965de081d2f716101c

Observation 00e68bf7-7e69-460b-8d0d-056198a81de3 · outbound

This paper cites Advances in Neural Information Processing Systems 37:1750–1810.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems 37:1750–1810

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.659874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.659874Z digest=sha256:d774b41e18192cfd5eb0000aedaf00e494f499d10df7b5548fd5c82a5fd4a8d8

Observation 68a30ee7-11c0-4080-b444-3ca23fa47e20 · outbound

This paper cites Thirty-seventh Conference on Neural Informa- tion Processing Systems.

Statistical and Algorithmic Foundations of Reinforcement Learning Thirty-seventh Conference on Neural Informa- tion Processing Systems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.662819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.662819Z digest=sha256:76a54951dddb16077a7a3474185b7a8e109a170859fcb7d59604c1ed6be32afb

Observation bd11732f-6a12-4c64-9a47-248a3901ee7c · outbound

This paper cites Algorithmic Learning Theory , 578–598 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning Algorithmic Learning Theory , 578–598 (PMLR)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.665426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.665426Z digest=sha256:413c8d1ec206b2f618a3d0cddf5735cde0f88ada9daa7af9348f371a4b2fd1fa

Observation adca8aff-9532-4ec8-85aa-12f2d1a0e8a3 · outbound

This paper cites Q-learning with UCB Exploration is Sample Efficient for Infinite-Horizon MDP.

Statistical and Algorithmic Foundations of Reinforcement Learning Q-learning with UCB Exploration is Sample Efficient for Infinite-Horizon MDP

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T16:15:16.203077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.669216Z digest=sha256:8f35bd224ee97f2e7641b7ed76127fd63c2404b0ab8eb8f10c105bbb03e43810

Observation c0d3bcea-a478-43e1-8754-535767a3c107 · outbound

This paper cites Bilinear Classes: A Structural Framework for Provable Generalization in RL.

Statistical and Algorithmic Foundations of Reinforcement Learning Bilinear Classes: A Structural Framework for Provable Generalization in RL

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.672091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.672091Z digest=sha256:be2f4b612e54d1f315becc62dddf50c778426cd5f437918cefef1708d82f90ef

Observation 120fdbc4-e8eb-46f1-aa2a-bac11f48c4c3 · outbound

This paper cites The Annals of Statistics 49(3):1378–1406.

Statistical and Algorithmic Foundations of Reinforcement Learning The Annals of Statistics 49(3):1378–1406

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.675019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.675019Z digest=sha256:7157d4199e10bcf1b8f9609f5ceb81ea9d664033f73b0b9cd98045d03f0f5544

Observation c68b7e45-697b-4470-a7a9-4fa2a735d3f6 · outbound

This paper cites Journal of machine learning Research 5(Dec):1–25.

Statistical and Algorithmic Foundations of Reinforcement Learning Journal of machine learning Research 5(Dec):1–25

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.677860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.677860Z digest=sha256:4654e6bfd5158b9d9f182ac48b7bed54c12b486b057cf5aead9271207336f3c8

Observation 0980d5a1-9f25-4414-b556-a09c96c08a38 · outbound

This paper cites International conference on machine learning, 1467–1476 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International conference on machine learning, 1467–1476 (PMLR)

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.680580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.680580Z digest=sha256:32440de8cd53695c9091204451590b7b244dff453b446470e651266a620aed89

Observation 87b205ca-c24e-4ca0-8376-572f95508cd7 · outbound

This paper cites The Statistical Complexity of Interactive Decision Making.

Statistical and Algorithmic Foundations of Reinforcement Learning The Statistical Complexity of Interactive Decision Making

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.683085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.683085Z digest=sha256:3624232a5f1490ddfbe377bf49431a0d8a21418f9fceb6f464275020711380a9

Observation ee3a6a52-4683-45a7-8b86-9a2b453bf786 · outbound

This paper cites Operations Research 71(6):2291–2306.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 71(6):2291–2306

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.685895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.685895Z digest=sha256:31372a68247276152695f76d24d45874e378415573b7641120c43f1bd8cc26f2

Observation 1833bfe8-ebd1-4a3b-a8c4-f0616d57827a · outbound

This paper cites Advances in neural information processing sys- tems 23:2613–2621.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing sys- tems 23:2613–2621

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.688671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.688671Z digest=sha256:3871cfd60bb97398974916cb4966d8bf9ce738427b05752a5123149671841904

Observation b0cd9721-5fbf-45a1-990b-56f3627f5469 · outbound

This paper cites International Conference on Machine Learning , 12790–12822 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 12790–12822 (PMLR)

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.691313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.691313Z digest=sha256:52874e1b229729db79f08062ba253ab51fba792178be71ed972e6e1666ddb56f

Observation a2356951-8c33-4ec8-8520-daec7a9f2595 · outbound

This paper cites Mathematics of Operations Research 30(2):257–280.

Statistical and Algorithmic Foundations of Reinforcement Learning Mathematics of Operations Research 30(2):257–280

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.693866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.693866Z digest=sha256:68033bcc3dc16d2e035a2919c71974095696f73f5a5db411b4058bc8e4d70934

Observation 3644e780-3a6e-426a-9fc4-c14cb33110c6 · outbound

This paper cites Journal of Machine Learning Research 11:1563–1600.

Statistical and Algorithmic Foundations of Reinforcement Learning Journal of Machine Learning Research 11:1563–1600

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.696679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.696679Z digest=sha256:2081b234a8b5647164406dfe53c68fccf915fa703bfcd47e5e1c303c9952e417

Observation c08b3e32-668a-4811-b53c-961dcb296400 · outbound

This paper cites Advances in neural information processing systems 36:80674–80689.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 36:80674–80689

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.699236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.699236Z digest=sha256:90391ee6375c1479132d817e8df53f147046f40eb7a24992a84866dafdfa5f15

Observation cbd95dce-b202-43c5-a955-c69a785c2fc8 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.702414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.702414Z digest=sha256:6b6635b38f7033f7548b530769c526573229062b6abde53bf62e188c500d1314

Observation 7c41e9f2-608d-4238-85d4-bb11ea4971dc · outbound

This paper cites International Conference on Machine Learning , 4870–4879 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 4870–4879 (PMLR)

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.704872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.704872Z digest=sha256:85e1405ef8592e34abb572e8e60eecd95b37f0d012488741bce9b7d0f4c8ecca

Observation 6429a0f9-983a-42cb-8f49-402fd2420247 · outbound

This paper cites Advances in neural information processing systems 34:13406–13418.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 34:13406–13418

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.707540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.707540Z digest=sha256:24d3dd5e3bdd08ca058ed5b0da6215e6f75007e736113d53c76c2812af21083a

Observation 33f1fb05-9c22-4701-925e-4a7118ceb632 · outbound

This paper cites Conference on Learning Theory , 2137–2143 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning Conference on Learning Theory , 2137–2143 (PMLR)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.710001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.710001Z digest=sha256:0e07fe8e76889b74e9cf3c4f7d158272c21c84719d0ae75024ecc2b1a16a32eb

Observation 677e47ea-b9fb-4f06-bf79-4917285a202e · outbound

This paper cites Truncated Variance Reduced Value Iteration.

Statistical and Algorithmic Foundations of Reinforcement Learning Truncated Variance Reduced Value Iteration

Reference 50

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T16:15:16.171403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.712851Z digest=sha256:f46fafece08c91ba6fb0e2c34308c74a4b81404216a7af38cff986bef768ff48

Observation f949fb1b-dc6e-4789-859f-63169b561092 · outbound

This paper cites Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality.

Statistical and Algorithmic Foundations of Reinforcement Learning Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T16:15:16.158454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.715810Z digest=sha256:a325e3b0820f40c027d66e320d5faabce59f71c0551d45af1711a8d344a47364

Observation 1e43d55b-b08e-401e-88d1-76b69a147cd2 · outbound

This paper cites International Conference on Machine Learning , 5055–5064 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 5055–5064 (PMLR)

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.718629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.718629Z digest=sha256:fb1b55e9bc030819cdf1449543120f20bf3d5ed1e4f91956698e6bd1a005ef65

Observation 5554fea2-2587-4baf-9d57-0f1080817287 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.721328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.721328Z digest=sha256:a372e4c54662a29261385f19f62f59f4da5752c71f86f7618ee4033fe71cecf5

Observation b28c4f96-27c9-4490-8466-9fcd0e7c932c · outbound

This paper cites Advances in neural information processing systems , 315–323.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems , 315–323

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.723876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.723876Z digest=sha256:cdf2e4884450c412c1e9c2ca7df4aba9998446bb7a2d3386be638ca397c7ee07

Observation f8510dad-89c5-495d-9802-f7cebb0342b4 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.726348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.726348Z digest=sha256:1e009a424a484a74544c8a7ac236fcb0679e382011458d3b6e6e36ac27a2ccee

Observation 58efa114-0175-4882-bc56-0db3b904a02a · outbound

This paper cites Advances in neural information process- ing systems , 1531–1538.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information process- ing systems , 1531–1538

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.729125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.729125Z digest=sha256:923c50960a4329ca55fcb478f39ba965aae7dabeb07c82dcd92005db741ab6d2

Observation db0b6f4c-6d30-438d-89ff-066c277f0666 · outbound

This paper cites Advances in neural information processing systems , 996–1002.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems , 996–1002

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.731783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.731783Z digest=sha256:1fbb0a735ec509645852e262bc9580e2dde15ae0b9cbe180c6ae4a2fac7abed0

Observation 7eb07294-e04f-419d-8d46-6c16c57f585d · outbound

This paper cites SIAM Journal on Math- ematics of Data Science 3(4):1013–1040.

Statistical and Algorithmic Foundations of Reinforcement Learning SIAM Journal on Math- ematics of Data Science 3(4):1013–1040

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.734332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.734332Z digest=sha256:ce7d59f26d258fb372f30ed5556a632f68ff47c0ce0ea7afdb244b16b4962983

Observation 6d59be80-373a-41e6-b7bf-2f842b4560c3 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.736951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.736951Z digest=sha256:10b9825d601e371e0266341447a0e3b8d74955be5ea5bf766b4aa1511fc524f5

Observation 231935af-ec32-4e6a-8ec6-05a9549586f4 · outbound

This paper cites IEEE Transactions on Automatic Control 27(1):137–146.

Statistical and Algorithmic Foundations of Reinforcement Learning IEEE Transactions on Automatic Control 27(1):137–146

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.739408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.739408Z digest=sha256:efc223217875c74d69e7005a7cc8815f9b2e78c7392463f4aeee0b1ceddf5be5

Observation d3cf1e35-b9f5-456b-801d-61730d3ed36b · outbound

This paper cites Advances in applied mathematics 6(1):4–22.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in applied mathematics 6(1):4–22

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.741940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.741940Z digest=sha256:63267f4c8350f7877eb316435baa0f4f6d07349716248e87f2d5dbc93f2e5f17

Observation 2d65c4ff-4309-4e1b-9d06-4e960aee8b1b · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.744401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.744401Z digest=sha256:01e37e9f216f88216f332ec27bfa91e2c9bd7dbeb445a8a98ceac9622b93df04

Observation 8b0b23a6-1835-4e94-9189-a25968bffa40 · outbound

This paper cites Reinforcement learning, 45–73 (Springer).

Statistical and Algorithmic Foundations of Reinforcement Learning Reinforcement learning, 45–73 (Springer)

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.747001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.747001Z digest=sha256:98d374ac1b48eb3c79e6697ec3d4981f158c6b69ddfed767bb22de55b519d882

Observation 1f0efe3a-3d01-4f72-a226-185d11a3fd50 · outbound

This paper cites International Con- ference on Algorithmic Learning Theory , 320–334 (Springer).

Statistical and Algorithmic Foundations of Reinforcement Learning International Con- ference on Algorithmic Learning Theory , 320–334 (Springer)

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.749643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.749643Z digest=sha256:57a621eccb967dc92aa4e8207b7e26e22ce2a7a45710f2d5640c9dc1ca5f41a9

Observation 7ae971a1-eb4e-4184-b555-a817de7eb4ba · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.752335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.752335Z digest=sha256:1e6486f914437fcb4c22c71599ce6bb8e7bca8d1fbf620c34e994eae59043b9e

Observation 29d3c5a0-3d60-47ff-9d3e-98def9604053 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Statistical and Algorithmic Foundations of Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.754970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.754970Z digest=sha256:12de3f5119ecfc2952f20f9822a3574be3150e35ff2f8b7c928cd56ef636298b

Observation 76e99e63-077b-4b42-a938-e929252dc9e3 · outbound

This paper cites Operations Research 72(1):222–236.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 72(1):222–236

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.757950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.757950Z digest=sha256:fa757d2f6b605a4ce2f873da4241a01e50deea0511b509ae54d66d145aa89096

Observation 365b5c58-e2d5-4bbc-bdd7-a4452d0be9f9 · outbound

This paper cites Advances in Neural Information Processing Systems , volume 35, 15353–15367.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems , volume 35, 15353–15367

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.760540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.760540Z digest=sha256:0aac6d611e6f0238fee58d9cc548e2023f80c7eeac9e3b70a2959ed88316e712

Observation 16202a1e-bc48-4efc-be55-a7f6c3b4a16b · outbound

This paper cites The Annals of Statistics 52(1):233–260.

Statistical and Algorithmic Foundations of Reinforcement Learning The Annals of Statistics 52(1):233–260

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.762992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.762992Z digest=sha256:5ef6ccd5b80cf2297cf3de1c6acb7176632d1878b06396698c10ebc8c4bf531c

Observation 43384b55-aee6-4b60-88a2-fa7b73a8955a · outbound

This paper cites Advances in Neural Information Processing Systems 34.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems 34

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.765580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.765580Z digest=sha256:b6d6d12cc680c2e1a6e6e87c2e311ed32e8885d4509e1b68b315b534f0b6f520

Observation 889b26fb-6ef1-4c17-9f19-be97bcce130d · outbound

This paper cites Mathematical Programming 1–96.

Statistical and Algorithmic Foundations of Reinforcement Learning Mathematical Programming 1–96

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.768171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.768171Z digest=sha256:7616b796b0af62240164dd5c005cd520e3893b22048857095d864a5ee62670a6

Observation 5afca019-fc67-43b4-846a-d16463f74b49 · outbound

This paper cites Operations Research 72(1):203–221.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 72(1):203–221

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.770763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.770763Z digest=sha256:0aecc1ff641c57795e36e5b4e5e576f385694d4f7f238ae5236b8e8e076c8dab

Observation 2676757a-1ae5-4152-89f1-aa9038b4a097 · outbound

This paper cites Advances in neural information processing systems 33:12861–12872.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in neural information processing systems 33:12861–12872

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.773621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.773621Z digest=sha256:c413a30392153949d7459cb7d508e2ebb00c73896332fea134459a71c0097d96

Observation 41f19b95-c669-42da-9205-b65746c4e5b2 · outbound

This paper cites IEEE Transactions on Information Theory 68(1):448–473.

Statistical and Algorithmic Foundations of Reinforcement Learning IEEE Transactions on Information Theory 68(1):448–473

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.776421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.776421Z digest=sha256:d5ca7fbd9e9f3b38606c6ac4585cf089c5f5e9989c13cd1ae25585610274ef64

Observation c894203d-872e-4178-a99e-5c11c8ca0cd6 · outbound

This paper cites The Thirty Seventh Annual Conference on Learning Theory , 3431–3436 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning The Thirty Seventh Annual Conference on Learning Theory , 3431–3436 (PMLR)

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.779018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.779018Z digest=sha256:e913b07a0e8d47945b63f3a7f17b37ca6f711309357b5eceaf0c7c5cf13f1f42

Observation 0b224d95-7049-47cb-a3d2-680bede3cb6a · outbound

This paper cites Advances in Neural Information Processing Systems 36.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems 36

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.781667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.781667Z digest=sha256:2da82931b7158095692e7c7db7e0c2d420c0805804e2ea713768365da9a16773

Observation c4385664-923b-41a0-9d53-f1c24af0a41c · outbound

This paper cites International Conference on Machine Learning , 6248–6258 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 6248–6258 (PMLR)

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.784390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.784390Z digest=sha256:fed3128d19e2f34d6f1d3b53b375e6321cf44647cebee2ae16495daaaabce905

Observation 6e9afc94-0d5e-4aca-86c8-c30dc179330a · outbound

This paper cites Games and economic behavior 10(1):6–38.

Statistical and Algorithmic Foundations of Reinforcement Learning Games and economic behavior 10(1):6–38

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.787425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.787425Z digest=sha256:dcd7a3e1edc171988e140d57782ca7b081e8bea996a581cc52c44ad6afa44c4a

Observation 86485cd9-1eb1-4e62-b2e9-a36ef0a60f4a · outbound

This paper cites International Conference on Machine Learning , 6820–6829 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning , 6820–6829 (PMLR)

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.790417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.790417Z digest=sha256:01b71bb500b818a87968ecf9bb95f7f6717427f2c364fc18ba170050bc1a5f4d

Observation cfeff9aa-4b04-4544-a3db-bba947385458 · outbound

This paper cites Conference on Learning Theory, 2947–2997 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning Conference on Learning Theory, 2947–2997 (PMLR)

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.793166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.793166Z digest=sha256:5a94df9d2c3679716489890a8006fb28b2cc6a5bdf07b8607922fe7340f2cc04

Observation f0fcbd82-6ce0-4076-9479-5f58eeac5fba · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.795858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.795858Z digest=sha256:af0b97d6f3e5c3cdaa1e5e9514031f1449841402adb76a56f40698375bbb4fa6

Observation d72d4ffb-933b-4980-9b1f-c809a5f249d4 · outbound

This paper cites Operations Research 53(5):780–798.

Statistical and Algorithmic Foundations of Reinforcement Learning Operations Research 53(5):780–798

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.798817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.798817Z digest=sha256:2771077fb033793eec78b77dcc3f8b572426aefa0b7c921f0fd784f4c9ac62e8

Observation 40774393-8e86-4690-b0fa-52b10eef868c · outbound

This paper cites (2022) Training language models to follow instructions with human feedback.

Statistical and Algorithmic Foundations of Reinforcement Learning (2022) Training language models to follow instructions with human feedback

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.801688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.801688Z digest=sha256:015c4a694a0d89a5cd333dbf25a0d0155abce9e5cc24b65c179e8f918d58a789

Observation 7f96ae58-f510-49b4-9af8-4c014d53ea7c · outbound

This paper cites International Conference on Artificial Intelligence and Statistics, 9582–9602 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Artificial Intelligence and Statistics, 9582–9602 (PMLR)

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.804449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.804449Z digest=sha256:816387c59e12b938ecb4f2e96edaeec4f52ba09caad0d9e2c07d4d295063a31b

Observation 480562cf-e498-4aed-b5a0-4e3156a11dba · outbound

This paper cites IEEE Transactions on Information Theory 67(1):566–585.

Statistical and Algorithmic Foundations of Reinforcement Learning IEEE Transactions on Information Theory 67(1):566–585

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.807408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.807408Z digest=sha256:36f74efe0b93d4dc518bbc25ec4347870d7e06074a453702780560930cb846f3

Observation 29a58c46-9992-482f-9d3c-f74b71cdbdf2 · outbound

This paper cites A Survey on Offline Reinforcement Learning: Taxonomy, Review, and Open Problems.

Statistical and Algorithmic Foundations of Reinforcement Learning A Survey on Offline Reinforcement Learning: Taxonomy, Review, and Open Problems

Reference 86

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T16:15:16.134078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.810157Z digest=sha256:67b3764984539ffd3d6d33d02dff702c57f5db116e98cb7ff369c9316e8e2322

Observation 2e51058d-276c-4991-9ef1-1241da4fe643 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.812922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.812922Z digest=sha256:c98f2a3d18b090d94e0715a22d87ebac35c6a899344b112053c7dac717f309cb

Observation 30e65d9d-c35e-48cd-a829-0a74a531f6cf · outbound

This paper cites Conference on Learning Theory 3185–3205.

Statistical and Algorithmic Foundations of Reinforcement Learning Conference on Learning Theory 3185–3205

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.816021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.816021Z digest=sha256:603dfbc8a83f7b8039cc633830d84b1708c5a0c5a69721f4f9502820f1ffe9d4

Observation 3b8033e1-1767-48d8-ac46-09c54e5dc3f4 · outbound

This paper cites Advances in Neural Information Processing Systems 36.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems 36

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.818702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.818702Z digest=sha256:78197ea8a2eb6e508f459525ab45dfa2f24631a87e2b18f26cbb2584444e8706

Observation b98d7279-ee7e-410b-b6eb-7c56223f2ad9 · outbound

This paper cites IEEE Transactions on Informa- tion Theory 68(12):8156–8196.

Statistical and Algorithmic Foundations of Reinforcement Learning IEEE Transactions on Informa- tion Theory 68(12):8156–8196

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.821727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.821727Z digest=sha256:1c07f4bf6557ca5e2dc4fd8f319381d87f83736973e74d08532e2cbc0427dd72

Observation 8c46f309-995d-4c84-9d8d-ee287d1eb483 · outbound

This paper cites Proceedings of the 35th International Conference on Neural Information Pro- cessing Systems, 15621–15634.

Statistical and Algorithmic Foundations of Reinforcement Learning Proceedings of the 35th International Conference on Neural Information Pro- cessing Systems, 15621–15634

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.824494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.824494Z digest=sha256:0e89cba9d9e4b45b2bf4fbdeeb818d0f2d7271356c82f7242e118f69d215ff57

Observation a6617798-7f47-455d-882a-528aaa128892 · outbound

This paper cites The Annals of Math- ematical Statistics 400–407.

Statistical and Algorithmic Foundations of Reinforcement Learning The Annals of Math- ematical Statistics 400–407

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.827171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.827171Z digest=sha256:d04be767637f9024c336fd2353a6adf63910d2086f012705f857f58d0dea93e2

Observation 011c3f51-53b2-446a-ac6a-2aad727abf49 · outbound

This paper cites The Thirty-eighth Annual Conference on Neural Information Processing Systems.

Statistical and Algorithmic Foundations of Reinforcement Learning The Thirty-eighth Annual Conference on Neural Information Processing Systems

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.867130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.830079Z digest=sha256:7db930356dd138384df7a2cb4cb3d882595f62c108875f5287a7adbea20febd8

Observation c2a0dbfe-341e-4840-8cde-895f6b06e181 · outbound

This paper cites The Thirty Seventh Annual Conference on Learning Theory , 4511–4547 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning The Thirty Seventh Annual Conference on Learning Theory , 4511–4547 (PMLR)

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.858474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.832787Z digest=sha256:ada537ff15b2553bd7b68e75b495f024b0340aa27d0211b8b9ca5c1373e34c23

Observation 71c57e02-a541-4377-b28d-d307580ff2a6 · outbound

This paper cites an unresolved cited work.

Statistical and Algorithmic Foundations of Reinforcement Learning Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:15:16.850078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.835923Z digest=sha256:842cf39e69527728bc471d4615a479f5f5c456f34fdff0877413c7b7230a1879

Observation d4ed827c-e1e5-4db4-8080-34e1ab6f1ca1 · outbound

This paper cites Journal of Machine Learning Research 25(200):1–91.

Statistical and Algorithmic Foundations of Reinforcement Learning Journal of Machine Learning Research 25(200):1–91

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.841650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.839243Z digest=sha256:b078f8adf0f93124c3359ddc005ee7dd0cd62bb8ccb89eef5f47e10f507c4531

Observation 9af1ae1b-f334-4b7a-99f9-71de1e318321 · outbound

This paper cites International Conference on Machine Learning (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning (PMLR)

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.832857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.842026Z digest=sha256:f3126e0ad1c3c0d8e508b9e9a5b74b3cf67555cf7cd964e9f62dd719de56f421

Observation 3af75634-3043-425f-a06d-6d45a12613eb · outbound

This paper cites International Conference on Machine Learning 19967–20025.

Statistical and Algorithmic Foundations of Reinforcement Learning International Conference on Machine Learning 19967–20025

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.824681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.844965Z digest=sha256:0bae5a171f0e0ad0db5c55a92eb0358bb5057e308aa67d55e0120e5663413869

Observation 0755f1ff-6864-4c27-8071-4821c2eb2d8e · outbound

This paper cites Advances in Neural Information Processing Systems 36.

Statistical and Algorithmic Foundations of Reinforcement Learning Advances in Neural Information Processing Systems 36

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.815809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.848230Z digest=sha256:4abfceb52caf23a06cc464bda0c985b605ba3e3a40284bffca44f2de68e2146c

Observation 5f5a39a2-8ae0-45d6-8cd3-5aaf5ba905ff · outbound

This paper cites International Confer- ence on Machine Learning , 44909–44959 (PMLR).

Statistical and Algorithmic Foundations of Reinforcement Learning International Confer- ence on Machine Learning , 44909–44959 (PMLR)

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:15:16.807059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:15:15.850895Z digest=sha256:553261d3ab5ea5c013c32ed9cc471ce9c23809d6f01c6623b34b33694b273b92

Pith citing papers

Observation b012529b-b82f-4ebe-bc86-7eedd4dd0722 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization Statistical and Algorithmic Foundations of Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:a419e0c8d15a0ed1f948667cbcdde51a545aee51800597c6a7a7340586deec42