Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:19:16.651877Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 100 of 109 outbound references and 0 inbound Pith citation observations for arXiv:2501.02089.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:19:16.651877Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 109 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8cd67070-287d-4075-be8a-2fa3de722aa9 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Im- proved algorithms for linear stochastic bandits
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1529a580-e3fc-49a1-886e-e55ef4ae0448 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Model-based reinforcement learning with a generative model is minimax op- timal
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d79c0843-8f33-40dc-b94f-c4a8d025be67 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Degenerate nonlinear programming with a quadratic growth condition
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4dce0c3-ce49-4107-bfd8-61e49fb6134b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Fitted q- iteration in continuous action-space mdps
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b85e5d2-4428-4f56-bda2-9ae3cc6a4408 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Learning the target network in function space
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95f88b01-9202-49e2-ad3c-3b1126b78f35 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Finite-time analysis of the multiarmed bandit problem, 2002
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bfac92d-a9d1-4350-9b10-538a891de260 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Minimax regret bounds for reinforcement learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 612dbc69-42f1-4026-aac5-f9cb2c4542aa · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Prov- ably efficient q-learning with low switching cost
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 391d16ad-2c33-4164-bce9-9a283ade1ad1 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Training a helpful and harmless assistant with reinforcement learning from human feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fed46ad-9f26-4721-a0fd-d6a4296b182e · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Dynamic programming
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d5bcf53-36f9-43ae-95c6-59bda6936c10 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Online learning with switching costs and other adaptive adversaries
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 579c1e47-5e27-49f6-b5c0-15309c726438 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Information-theoretic considera- tions in batch reinforcement learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e30b8fe-222f-4a3d-8460-5f0773f20b1e · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Deep reinforcement learning from human preferences
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b77f5264-9dea-4525-966d-53eb0a2c6bb8 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30728537-2956-4a7d-9cae-85becbeb2052 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Minimax-optimal off- policy evaluation with linear function approximation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f664f1a-61d3-4d8b-b260-dafc2f2b3c11 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Doubly Robust Policy Evaluation and Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff4366c-7ce2-4045-b441-874bd0c7ad4b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Tree-based batch mode reinforcement learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e421dd0-e1f5-4858-aff6-0396b456a2f3 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Discovering faster matrix multipli- cation algorithms with reinforcement learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1547a81a-a1db-4e9e-8690-9deba09ac06b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Theory of statistical estimation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23adf15-01b8-408c-ba93-c0a3705e9685 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures A Provably Efficient Algorithm for Linear Markov Decision Process with Low Switching Cost
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6db73a91-af48-4e56-80e4-676867e272e1 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Batched multi-armed bandits problem
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cac3d414-6306-424a-b5bb-2427cbdfb052 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Off-policy deep rein- forcement learning by bootstrapping the covariate shift
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 961f3efe-270f-40b7-9217-70dd5c1cf63b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Minimax pac bounds on the sample complexity of rein- forcement learning with a generative model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 689b4757-97cd-4604-b07c-f3f151f94402 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Maxmin expected utility with non-unique prior
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b6a0a7d-db16-4d7a-98c6-b9ab21630074 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Approximate solutions to Markov decision processes
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fe80bfc-4112-4f5f-b37b-55194084191d · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Networkgym: Reinforcement learn- ing environments for multi-access traffic management in net- work simulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ca303f-2507-4a8c-a827-34d6d588f2ca · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Consistent on-line off-policy evaluation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05eb0f54-0dd3-451b-b852-3dffe9027cc4 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Bootstrapping fitted q-evaluation for off- policy inference
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 667d8a66-1ef5-4361-8463-37b978db587a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Effi- cient estimation of average treatment effects using the estimated propensity score
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d26e45f8-545a-4681-901f-2b34b8f84dab · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures A generalization of sampling without replacement from a finite universe
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0b9d2f9-6be9-4d6c-b1c6-7370f8df9449 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Towards deployment-efficient reinforcement learning: Lower bound and optimality
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 043356c3-92bf-4056-b959-a07c8de46a47 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Aleatoric and epis- temic uncertainty in machine learning: An introduction to con- cepts and methods
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c4b3a65-6f0c-404d-aa1f-1ff225bc51bd · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Doubly robust off-policy value eval- uation for reinforcement learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e6085e59-902e-41c4-91d9-e09f3fc8ff07 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Offline reinforcement learning in large state spaces: Algorithms and guarantees
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f52c8254-9bed-41b8-9dff-eb535d10df04 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Is pessimism provably efficient for offline rl? In International Conference on Machine Learning, pages 5084–5096
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b8a3ab84-a748-4bcd-b827-85ff45cac308 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Double reinforcement learning for efficient off-policy evaluation in markov decision processes
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5a83a642-146b-4251-9a99-64b8b41d7be2 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Efficiently breaking the curse of horizon in off-policy evaluation with double reinforce- ment learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bafe5139-e632-4f6d-8cb9-5edc5238010c · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Near-optimal reinforce- ment learning in polynomial time
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6226483e-26dc-4d44-9f62-557c512b6647 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Introduction to empirical processes and semiparametric inference, volume 61
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8641d3d2-57e5-4f99-8e8a-5d4477510b90 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Offline reinforcement learning with fisher divergence critic regularization
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0663e18c-acf2-4cba-9ed9-b16a229433a0 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Conservative q-learning for offline reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3aa17862-1c9b-4384-92c8-b34eecdbfe0a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Minimax theory
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a2ba5fa-fec8-4916-8dc0-a6bd09c3fd44 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Batch policy learning under constraints
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 53fbf41f-b39e-4273-baa9-63ed9b715275 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215ce0da-c7fa-4b64-8547-77caf4e57389 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Breaking the sample size barrier in model-based reinforcement learning with a generative model
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 49255547-06f2-4497-9765-454caa65547e · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Offline reinforcement learning with closed-form policy improvement operators
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f09321f7-cc91-43c8-804d-e6d4873c4936 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Unbi- ased offline evaluation of contextual-bandit-based news article recommendation algorithms
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a16b704-7ff1-4654-80e9-82d0632306c6 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Monte Carlo strategies in scientific computing, volume 10
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 78954958-2c3a-4b48-8a42-23141e5ff526 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Breaking the curse of horizon: Infinite-horizon off-policy es- timation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 57a75e49-66aa-434d-9959-297a5db9a7af · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Off-policy policy gradient with stationary distribution cor- rection
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ae97e546-50b3-4c9b-bf78-01671a0aa7ee · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Provably good batch off-policy reinforcement learning without great exploration
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 405675f4-ba25-4549-8375-532ad9945af0 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Mildly conservative q-learning for offline reinforcement learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9e7b0ac0-cd9f-41d8-acd5-3af09958e87e · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures the distribution-norm to the res- cue
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 54ed687c-d2c2-46b3-a236-996f74795263 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Faster sorting algorithms discovered using deep reinforcement learn- ing
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1e0d4ece-f616-4e4f-81d8-bdb7858b2412 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Neural adaptive video streaming with pensieve
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8b2e305d-de76-487c-835f-3755f4a0be3a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Deployment-efficient reinforce- ment learning via model-based offline optimization
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cb40f6e4-1f27-4e42-bff5-f758f1b319b9 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Dependent central limit theorems and in- variance principles
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 58ccbf41-73d4-412d-89f5-14b877a01c8b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Variance-aware off-policy evaluation with linear function ap- proximation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e34d18aa-ef27-4d2a-9fb0-74054f8e5149 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Human-level control through deep reinforcement learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dd6bb0c5-0cfa-427b-b9ca-0fed1a73afff · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Bootstrapping: A nonparametric approach to statistical infer- ence
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e59f5fa6-37ca-4640-8f65-b706af237290 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Finite-time bounds for fit- ted value iteration
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 193c6ec1-b31e-4457-95f6-cac07abce26a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Marginal mean models for dynamic regimes
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3ccf2f67-d769-44e9-9bce-ee8a2068faaf · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Dualdice: Behavior-agnostic estimation of discounted stationary distribu- tion corrections
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 95de1a03-1199-4c28-ad5b-2985258fee37 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0d4c04b-e228-4175-ae87-a8383b436721 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Optimal medication dosing from suboptimal clinical ex- amples: A deep reinforcement learning approach
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c3585e0a-0abb-469f-9e0b-71f14dfd013e · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures On sample-efficient of- fline reinforcement learning: Data diversity, posterior sampling and beyond
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 19b323a4-cbe9-4848-895f-91a27aa5631a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures On instance-dependent bounds for offline reinforcement learning with linear function approximation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2d4b2c17-a3f3-47bd-bd3d-27744fb6f1b5 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2e04c9c0-c068-49f1-81b0-e36439b6e035 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Batched bandit problems
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4c44f74f-5eb9-4085-b8d9-9d5dd9f2e6d4 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Approximate Dynamic Programming: Solv- ing the curses of dimensionality , volume 703
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e5b586e5-8f70-4316-b061-bf6d2bf8990a · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Eligibility traces for off-policy policy evalua- tion
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 45d70d9d-3624-4ccc-8f27-5dd679f15df5 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Markov decision processes
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 563ed0ae-d2b1-41ed-91ff-088985f8c4ae · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Near-optimal deployment effi- ciency in reward-free reinforcement learning with linear func- tion approximation
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5cc730c1-b9ae-4717-bff6-d5aed1d48b13 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Sample- efficient reinforcement learning with loglog (t) switching cost
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 90471ee9-5843-4fbd-be07-5c0ef964fc56 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Logarithmic switch- ing cost in reinforcement learning beyond linear mdps
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d74fba3d-d6a5-40ca-bb2e-ac5dc47a8672 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Bridging offline reinforcement learning and im- itation learning: A tale of pessimism
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 08b9b6dd-f4fb-4a25-bcee-9eba6b35c3e6 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Nearly horizon-free offline reinforcement learning
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73f5d59e-d8e3-46ff-8735-08a530c69867 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Mastering the game of go with deep neural networks and tree search
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 60a06f68-e025-467a-9587-758852db3281 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Mastering the game of go without human knowledge
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa648aef-6fb2-4b3c-9e69-3af10e8a56ca · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Learning to summarize with human feed- back
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d0e89996-56c5-4de9-b783-0482330b2bde · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Reinforcement learn- ing: An introduction
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5c4d9419-5927-4a22-adc4-ac8db1fc9fd2 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Finite time bounds for sampling based fitted value iteration
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1f620bd3-35e8-4b44-a8e0-c92594300be9 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Semiparametric theory and missing data, volume 4
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5253eba6-d85a-49eb-a3e9-63f1629d1595 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Minimax weight and q-function learning for off-policy evaluation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fdef9586-60e1-47ab-aaa9-a545b885a559 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Asymptotic statistics, volume 3
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f87c33c9-fd40-4ca5-8ba6-c586cc08db69 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures High-dimensional statistics: A non- asymptotic viewpoint, volume 48
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7b684ee2-762a-4cab-bb9b-37310552c199 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Provably efficient reinforcement learning with linear function approxi- mation under adaptivity constraints
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d95b0957-9ce6-4423-818d-b7197348b8e4 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures On gap-dependent bounds for offline reinforcement learning
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0646dc31-11d4-48ce-b901-1af3c63df95b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Opti- mal and adaptive off-policy evaluation in contextual bandits
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 17268b18-fa42-4ba1-9a94-47bdce98e4c1 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Behavior Regularized Offline Reinforcement Learning
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c6f5143-4834-4b7f-b907-f772f2121fc5 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures On the optimality of batch policy optimization algorithms
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9757f4e0-0dbe-4062-bba3-b3cfcc744b21 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Q* approximation schemes for batch reinforcement learning: A theoretical comparison
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 189aff4c-fe89-4510-a5a7-af8d297a211c · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Towards optimal off-policy evaluation for reinforcement learning with marginal- ized importance sampling
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation da2895fe-2f71-49aa-bfa3-6016f79f0d28 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Bellman-consistent pessimism for offline rein- forcement learning
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 19817067-7e7a-4a91-9f41-05d340030dd3 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Policy finetuning: Bridging sample-efficient offline and online reinforcement learning
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6064c2ed-6b28-489a-8a2b-74a6bfba6e3b · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Nearly Minimax Optimal Offline Reinforcement Learning with Linear Function Approximation: Single-Agent MDP and Markov Game
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dfea6f6-3eb1-4951-bbdd-83da7727c7d5 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Asymptotically efficient off- policy evaluation for tabular reinforcement learning
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 80472f83-d677-47e9-aeff-1e7a3590bdcf · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Towards instance-optimal of- fline reinforcement learning with pessimism
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cc4bef7c-c292-402e-8207-a764985f2dcf · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Near-optimal prov- able uniform convergence in offline policy evaluation for rein- forcement learning
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5397d07a-78d9-4e51-baa1-2aa64fa26568 · outbound
On the Statistical Complexity for Offline and Low-Adaptive Reinforcement Learning with Structures Near-optimal of- fline reinforcement learning via double variance reduction
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
No inbound Pith citation observations are available.