Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:35:27.824202Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2508.08436.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:35:27.824202Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 60bb9e6b-52c1-4116-9982-3b30e9f26cf4 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Normal approximation for stochastic gradient descent via non-asymptotic rates of martingale clt
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3199ac5b-498f-4ef3-9568-7f0e4bdfbafb · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba82c5a5-668d-4e0b-9964-43deed480fbc · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Logarithmic regret for episodic continuous-time linear-quadratic reinforcement learning over a finite-time horizon
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d460d6e4-d5e4-47c8-bee7-cdd8a10eb7c0 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control with the stochastic approximation algorithm: Geometry and convergence
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aaad1fa6-b5a3-43de-927a-534f9dcfc526 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic programming and optimal control: Volume I , volume 4
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9b498ff2-d462-4347-bcde-18f845d587a2 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning applied to linear quadratic regulation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1157fc72-6af4-465d-b37b-4467c5d41ba4 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9542fdcd-b05b-4aa5-b292-357ad7a7166d · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for online decision making via stochastic gradient descent
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39472f0f-c982-4e40-b087-97cc71a7fa31 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for model parameters in stochastic gradient descent
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2c11bd4-9725-4966-ba3b-3c76681f879b · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust inference via multiplier bootstrap
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 618178cc-5fd5-4d0c-90fc-dc6a9eebd61d · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the sample complexity of the linear quadratic regulator
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2d1892e-a99d-49fe-a8bd-3a3e62f4c9fe · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61a1fb16-e1db-4c3e-a0b2-89e0f6070336 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of real-world reinforcement learning: definitions, benchmarks and analysis
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16510445-8a5f-40b8-98ab-4543b24c3fb7 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the convergence theory of debiased model-agnostic meta-reinforcement learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d085cb1-e113-402c-9a05-f1995d8cebc5 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap confidence intervals for the stochastic gradient descent estimator
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43d704f2-298b-4d3d-a850-d1d5cad672cc · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Global convergence of policy gradient methods for the linear quadratic regulator
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 784d7f5b-b5ee-4f7f-9fc0-2518b2c347cc · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56470e8b-ced4-404f-8470-f61f57a7ba36 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e84d419-1be6-43ef-a171-90223ace1d2d · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Policy gradient methods for the noisy linear quadratic regulator over a finite horizon
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 230fdaaf-92fe-43e4-b115-fa6a57d1d6fb · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Bootstrapping upper confidence bound
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1824c0d6-9934-4a8b-875a-ce502789a0ba · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Model-based or model-free, a review of approaches in reinforcement learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3bd7a1f-f21a-4334-be3d-23566858cd6c · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Improved zeroth-order variance reduced algorithms and analysis for nonconvex optimization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66e8339b-58a2-4b8c-ba4f-0dbee8facd9e · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On incomplete learning and certainty-equivalence control
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc17931f-76a1-4e95-8317-0c21a9ca6694 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Iterated least squares in multiperiod control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2257be21-7815-492a-a897-19265deda77c · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Fast inference for quantile regression with tens of millions of observations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 589ef146-8f4a-4082-a5fc-dee4d511edaf · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical Estimation and Inference via Local SGD in Federated Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d1b0510-8ed4-43e9-87d0-08bd5e84a753 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Unifying offline causal inference and online bandit learning for data driven decision
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d034d63-7a58-4444-8405-6dffb8f0d0c5 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A generalized reinforcement-learning model: Convergence and applications
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97fa945a-a8a3-4591-bb2f-d8e8db676be7 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A review of uncertainty for deep reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d0b37d9-9973-49f8-8287-d6ee78c67900 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gradient estimation in model-based reinforcement learning: a study on linear quadratic environments
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f34550ab-8f43-4611-bf02-7b9c047cc283 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Certainty equivalence is efficient for linear quadratic control
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63d803c6-3821-42da-8ab8-8c9b79cc96e4 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A linear quadratic regulator based speed control for remote-controlled racing cars
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee475c6d-d57e-4554-b11f-910186482ddd · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Black-box generalization: Stability of zeroth-order learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a09a9f0a-2c0a-41fd-8c4a-7349123df957 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Acceleration of stochastic approximation by averaging
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73f85c50-cd5d-4224-b957-f177ab39d3e0 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Temporal Difference Models: Model-Free Deep RL for Model-Based Control
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70c83ace-2b1b-4688-a7de-a53ab20b67fd · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap inference for policy evaluation in reinforcement learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a0be36e-29f1-4e77-8f9e-722e063a427e · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon The unintended consequences of discount regularization: Improving regularization in certainty equivalence reinforcement learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33bd0d6e-8c77-488e-8e73-9b1ae9647dd4 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b898f0e-802a-4fc5-a8c1-7309592af786 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Efficient estimations from a slowly convergent robbins-monro process
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16618310-e5ab-42c0-a114-9808f6df6232 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gaussian approximation and multiplier bootstrap for polyak-ruppert averaged linear stochastic approximation with applications to td learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9853206d-8b5a-4474-862e-f13eda4574c2 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic causal effects evaluation in a/b testing with a reinforcement learning framework
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f24ae1fc-815c-4c3c-a630-755deec6d106 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Learning optimal controllers by policy gradient: Global optimality via convex parameterization
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82b60fb8-3072-4679-a76d-90a8211c62c1 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec2992d7-cfff-4afc-8363-59c56f6cce3f · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust exploration in linear quadratic reinforcement learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b34a2f0-c3bf-415c-bbaf-e3679dd4cd02 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Distributed lqr design for identical dynamically coupled systems: Application to load frequency control of multi-area power grid
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9516177-5a22-4215-8422-76abd2bc77a3 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon First-order regret in reinforcement learning with linear function approximation: A robust estimation approach
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0771b835-b586-4bf6-bf9a-a262f3fbfb0f · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Residual Bootstrap Exploration for Bandit Algorithms
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0e2f6e8-73f4-4fb0-a457-ff4d79088e7b · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Exact asymptotics for linear quadratic adaptive control
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8503830b-2585-4327-9b3c-0e64e498b5cf · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Continuous-time mean--variance portfolio selection: A reinforcement learning framework
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2331a859-145c-48a3-9385-e20128b4e3ea · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Stochastic zeroth-order optimization in high dimensions
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adc91b51-daa4-48e6-b62d-13db23fe28cf · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Leveraging linear quadratic regulator cost and energy consumption for ultrareliable and low-latency iot control systems
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1e7ea49-2d11-44bf-b1b8-8a2588cd0fe3 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1456cc1b-6615-44cd-b6f3-899786ffcef2 · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online covariance matrix estimation in stochastic gradient descent
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a418e2f7-6b01-417e-86f2-07c3a8339bae · outbound
Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Uncertainty quantification and exploration for reinforcement learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.