Pith. sign in

Paper Citation Record · LEDGER

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon

As of 7 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2508.08436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.08436 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:35:27.824202Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy47
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60bb9e6b-52c1-4116-9982-3b30e9f26cf4 · outbound

This paper cites Normal approximation for stochastic gradient descent via non-asymptotic rates of martingale clt.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Normal approximation for stochastic gradient descent via non-asymptotic rates of martingale clt

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.674006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.798849Z digest=sha256:3bda958bde8bd14544ce398106ff0e5dad8c60ee8fa494a86cb83267b27a2c41

Observation 3199ac5b-498f-4ef3-9568-7f0e4bdfbafb · outbound

This paper cites Adaptive control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.464013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.844321Z digest=sha256:d2f91f593b52b601bc12290bbbc1b95ccc15650818863d2f6314d69e97114c70

Observation ba82c5a5-668d-4e0b-9964-43deed480fbc · outbound

This paper cites Logarithmic regret for episodic continuous-time linear-quadratic reinforcement learning over a finite-time horizon.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Logarithmic regret for episodic continuous-time linear-quadratic reinforcement learning over a finite-time horizon

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.327979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.885776Z digest=sha256:4438452a7df5d44a03abb6428fd8cf5f02fb4bc2e74c3361a8bcb6be252ffe85

Observation d460d6e4-d5e4-47c8-bee7-cdd8a10eb7c0 · outbound

This paper cites Adaptive control with the stochastic approximation algorithm: Geometry and convergence.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control with the stochastic approximation algorithm: Geometry and convergence

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.121599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.974853Z digest=sha256:67ed3b129d065c389d817c2ba36a0746ece0767c9c6298c38d4da15469c45297

Observation aaad1fa6-b5a3-43de-927a-534f9dcfc526 · outbound

This paper cites Dynamic programming and optimal control: Volume I , volume 4.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic programming and optimal control: Volume I , volume 4

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.925538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.061034Z digest=sha256:6fe28471c6d4b84530e58661f7b9ece328a715281dd436a8672d3b174574e062

Observation 9b498ff2-d462-4347-bcde-18f845d587a2 · outbound

This paper cites Reinforcement learning applied to linear quadratic regulation.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning applied to linear quadratic regulation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.688872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.177822Z digest=sha256:1286c11822620352cd793100b61eb76941f402e034015944577f5ada6a5d272f

Observation 1157fc72-6af4-465d-b37b-4467c5d41ba4 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:24.252521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:24.252521Z digest=sha256:0a5aa8a91a0a67e4b802d7500d5fa4bb2b4ae335fd74b25758892cd93a8cd0bd

Observation 9542fdcd-b05b-4aa5-b292-357ad7a7166d · outbound

This paper cites Statistical inference for online decision making via stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for online decision making via stochastic gradient descent

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.529353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.312545Z digest=sha256:4c34ed935c8f94023b28f2ab0f1f1a4f2da6e4dc445e0fd0349de35153d2293a

Observation 39472f0f-c982-4e40-b087-97cc71a7fa31 · outbound

This paper cites Statistical inference for model parameters in stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for model parameters in stochastic gradient descent

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.356129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.351893Z digest=sha256:78e778a7c5ddbc0fb0094363b3fc4d366d5b2f7cb0e4dd1420484dba214a1e9c

Observation c2c11bd4-9725-4966-ba3b-3c76681f879b · outbound

This paper cites Robust inference via multiplier bootstrap.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust inference via multiplier bootstrap

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.130232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.431215Z digest=sha256:7269c8ab9cc0de4681d3daac3b91e285f525cffc161d0dc59882cc2f9d6a1739

Observation 618178cc-5fd5-4d0c-90fc-dc6a9eebd61d · outbound

This paper cites On the sample complexity of the linear quadratic regulator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the sample complexity of the linear quadratic regulator

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:24.507547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:24.507547Z digest=sha256:76fd97332fd217823f597401302c7e01a8868bd573e5a6418b73534aeabcc631

Observation a2d1892e-a99d-49fe-a8bd-3a3e62f4c9fe · outbound

This paper cites Challenges of reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.966921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.579776Z digest=sha256:a1bf54e9dafeb15c9e5ecfc23947228fbfe9364baf2a0cc222bb10ee1fa51f98

Observation 61a1fb16-e1db-4c3e-a0b2-89e0f6070336 · outbound

This paper cites Challenges of real-world reinforcement learning: definitions, benchmarks and analysis.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of real-world reinforcement learning: definitions, benchmarks and analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.812742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.650888Z digest=sha256:a0071720b5c9ad75bbb1864050b682bc0ab8ade7174928bf7a9975e0bb5c2c0f

Observation 16510445-8a5f-40b8-98ab-4543b24c3fb7 · outbound

This paper cites On the convergence theory of debiased model-agnostic meta-reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the convergence theory of debiased model-agnostic meta-reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.551613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.720287Z digest=sha256:1e43569c3fd0973f5bb4095ba1c7beb3fde7511bc56972f5758409bc296ed357

Observation 5d085cb1-e113-402c-9a05-f1995d8cebc5 · outbound

This paper cites Online bootstrap confidence intervals for the stochastic gradient descent estimator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap confidence intervals for the stochastic gradient descent estimator

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.253477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.763782Z digest=sha256:52577f8fa985ab89dd5e63ab6873fed4b46dde2d80675936e6f9d6f526bb9974

Observation 43d704f2-298b-4d3d-a850-d1d5cad672cc · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Global convergence of policy gradient methods for the linear quadratic regulator

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.063605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.816293Z digest=sha256:5cbbe357d0ae9fb9f08a1579c9526d96baf4dd9beaf345cd46e11f45eccfc02d

Observation 784d7f5b-b5ee-4f7f-9fc0-2518b2c347cc · outbound

This paper cites Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.832945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.894628Z digest=sha256:667928efd696b89eec4a70affbeb8ba7162fb8a97f4d54c2302d8a380784dd25

Observation 56470e8b-ced4-404f-8470-f61f57a7ba36 · outbound

This paper cites Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.518251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.975470Z digest=sha256:b910307b983ede218304a48045c58f79321aa051ff3fb7f695f5170bb7dad226

Observation 5e84d419-1be6-43ef-a171-90223ace1d2d · outbound

This paper cites Policy gradient methods for the noisy linear quadratic regulator over a finite horizon.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Policy gradient methods for the noisy linear quadratic regulator over a finite horizon

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.307106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.083912Z digest=sha256:881ead5f95fb2001e2f9bafaf23c546e37dbb05c9af2d0c8463f797c7c9526cf

Observation 230fdaaf-92fe-43e4-b115-fa6a57d1d6fb · outbound

This paper cites Bootstrapping upper confidence bound.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Bootstrapping upper confidence bound

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.045307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.186107Z digest=sha256:c734ac5df55b23aaa801d507cfcd5d980b398fae91dd2f106ffb5de90524e2ae

Observation 1824c0d6-9934-4a8b-875a-ce502789a0ba · outbound

This paper cites Model-based or model-free, a review of approaches in reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Model-based or model-free, a review of approaches in reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.809383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.271012Z digest=sha256:0521d6e669d0e416b21d78d4a747405ab891956f6709a2141bb1fa0ee5e849c1

Observation d3bd7a1f-f21a-4334-be3d-23566858cd6c · outbound

This paper cites Improved zeroth-order variance reduced algorithms and analysis for nonconvex optimization.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Improved zeroth-order variance reduced algorithms and analysis for nonconvex optimization

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.496520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.382262Z digest=sha256:1ff6139f159925b8feabaf9f62b991d452c758b75f88029873d6bd5679915701

Observation 66e8339b-58a2-4b8c-ba4f-0dbee8facd9e · outbound

This paper cites On incomplete learning and certainty-equivalence control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On incomplete learning and certainty-equivalence control

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.213608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.472692Z digest=sha256:3f8e764d94115f938d21d9f738b8b31fd58f949c3fd12cc00ee9f89c7c96599f

Observation bc17931f-76a1-4e95-8317-0c21a9ca6694 · outbound

This paper cites Iterated least squares in multiperiod control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Iterated least squares in multiperiod control

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.019692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.550290Z digest=sha256:d7c996b4025eda54bd856a6a0aa0a275a09f230c7b28147188ddddb37127c79a

Observation 2257be21-7815-492a-a897-19265deda77c · outbound

This paper cites Fast inference for quantile regression with tens of millions of observations.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Fast inference for quantile regression with tens of millions of observations

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.868085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.619163Z digest=sha256:4132458bd29091a51d647006e217cabe84462d828ee6840c461903384230ca73

Observation 589ef146-8f4a-4082-a5fc-dee4d511edaf · outbound

This paper cites Statistical Estimation and Inference via Local SGD in Federated Learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical Estimation and Inference via Local SGD in Federated Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:25.666238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:25.666238Z digest=sha256:5cb3fbc3e8bfdb0fb61c1ddf28a3ca517cdf68cf023520888550ff4ac71d1957

Observation 3d1b0510-8ed4-43e9-87d0-08bd5e84a753 · outbound

This paper cites Unifying offline causal inference and online bandit learning for data driven decision.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Unifying offline causal inference and online bandit learning for data driven decision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.661794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.727112Z digest=sha256:50004d1c40de24df41dbddbfb40d9e553c3625f477f4a75b7b57279cf9792d5d

Observation 4d034d63-7a58-4444-8405-6dffb8f0d0c5 · outbound

This paper cites A generalized reinforcement-learning model: Convergence and applications.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A generalized reinforcement-learning model: Convergence and applications

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.373166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.776992Z digest=sha256:6024d0c90083b479bc08e8673a979723b9795a4a13e9f57f590326e1f00be4d0

Observation 97fa945a-a8a3-4591-bb2f-d8e8db676be7 · outbound

This paper cites A review of uncertainty for deep reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A review of uncertainty for deep reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.091172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.853537Z digest=sha256:6bf893535b225dec1e76362f42993e7373e1046779a824fab8c70905196f599d

Observation 3d0b37d9-9973-49f8-8287-d6ee78c67900 · outbound

This paper cites Gradient estimation in model-based reinforcement learning: a study on linear quadratic environments.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gradient estimation in model-based reinforcement learning: a study on linear quadratic environments

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.798966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.986287Z digest=sha256:053aa22bfd98d0c56ec116a895958b82e42e696fc2f95e1e713b224fb15f3f77

Observation f34550ab-8f43-4611-bf02-7b9c047cc283 · outbound

This paper cites Certainty equivalence is efficient for linear quadratic control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Certainty equivalence is efficient for linear quadratic control

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.689389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.049926Z digest=sha256:e19c9cbcee5235653775856fa62508d09649d9f92478073614465315a1d81596

Observation 63d803c6-3821-42da-8ab8-8c9b79cc96e4 · outbound

This paper cites A linear quadratic regulator based speed control for remote-controlled racing cars.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A linear quadratic regulator based speed control for remote-controlled racing cars

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.563763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.115250Z digest=sha256:8f07651a5b2ae5d9c8873435a204ac7282579cb784618c86863e17c7d30f8f7d

Observation ee475c6d-d57e-4554-b11f-910186482ddd · outbound

This paper cites Black-box generalization: Stability of zeroth-order learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Black-box generalization: Stability of zeroth-order learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.405457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.193252Z digest=sha256:c0d820157f752ff50732a265eacaeb121729f1cf87eca3042eb08919b970788f

Observation a09a9f0a-2c0a-41fd-8c4a-7349123df957 · outbound

This paper cites Acceleration of stochastic approximation by averaging.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Acceleration of stochastic approximation by averaging

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.283756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.241782Z digest=sha256:8e531eec14bc857579ba9767c0beb0bf8dda35382e187bdaf04cb4f253654a5d

Observation 73f85c50-cd5d-4224-b957-f177ab39d3e0 · outbound

This paper cites Temporal Difference Models: Model-Free Deep RL for Model-Based Control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Temporal Difference Models: Model-Free Deep RL for Model-Based Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:26.295386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:26.295386Z digest=sha256:969317dbcef5feaa81aec2e9e35a763dbe9b1801982f9353ca29fe678b1abd50

Observation 70c83ace-2b1b-4688-a7de-a53ab20b67fd · outbound

This paper cites Online bootstrap inference for policy evaluation in reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap inference for policy evaluation in reinforcement learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.143073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.387339Z digest=sha256:53eaaf5d1849701a7f899c8c265bc3c4b0ecfd7bfc941b09034e04443e2dc842

Observation 0a0be36e-29f1-4e77-8f9e-722e063a427e · outbound

This paper cites The unintended consequences of discount regularization: Improving regularization in certainty equivalence reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon The unintended consequences of discount regularization: Improving regularization in certainty equivalence reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.983187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.452152Z digest=sha256:78c473783ff10b08679175c73c024fe86a3e25e56a85c99fdf0834fea225b1cc

Observation 33bd0d6e-8c77-488e-8e73-9b1ae9647dd4 · outbound

This paper cites Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:35:28.059162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.554802Z digest=sha256:7f543d68b858d11202b167bffd48a70c9967cd455579dd4330faa5e3f361eed5

Observation 5b898f0e-802a-4fc5-a8c1-7309592af786 · outbound

This paper cites Efficient estimations from a slowly convergent robbins-monro process.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Efficient estimations from a slowly convergent robbins-monro process

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:26.641853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:26.641853Z digest=sha256:9d619ead10b552e572ac24691ed22647d90bd3cca9d54fa95d474b7c0e189746

Observation 16618310-e5ab-42c0-a114-9808f6df6232 · outbound

This paper cites Gaussian approximation and multiplier bootstrap for polyak-ruppert averaged linear stochastic approximation with applications to td learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gaussian approximation and multiplier bootstrap for polyak-ruppert averaged linear stochastic approximation with applications to td learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.854106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.704743Z digest=sha256:6c58c19ccc6daea13948f1b4e610b9b9a71bdb53d4ec89a8436e3976fc68992e

Observation 9853206d-8b5a-4474-862e-f13eda4574c2 · outbound

This paper cites Dynamic causal effects evaluation in a/b testing with a reinforcement learning framework.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic causal effects evaluation in a/b testing with a reinforcement learning framework

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.720305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.808715Z digest=sha256:7edc897645de63e6dd6d9460cbae17411b5a0f47f1ec39113ce65435a597af61

Observation f24ae1fc-815c-4c3c-a630-755deec6d106 · outbound

This paper cites Learning optimal controllers by policy gradient: Global optimality via convex parameterization.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Learning optimal controllers by policy gradient: Global optimality via convex parameterization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.598742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.899905Z digest=sha256:615b2850ba8237981eb257369aba5e70ea86264c4a90506075e317b9d1e1b568

Observation 82b60fb8-3072-4679-a76d-90a8211c62c1 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.477974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.983368Z digest=sha256:1b4c658f9e7fc94b08419dd77647cd506e75a8314aeafcc2a14797bc98c19843

Observation ec2992d7-cfff-4afc-8363-59c56f6cce3f · outbound

This paper cites Robust exploration in linear quadratic reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust exploration in linear quadratic reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.368728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.100028Z digest=sha256:7c24f690f66bcfeec97bdbb760af02bc9626ab0072e14b94ae3d1934be4592af

Observation 1b34a2f0-c3bf-415c-bbaf-e3679dd4cd02 · outbound

This paper cites Distributed lqr design for identical dynamically coupled systems: Application to load frequency control of multi-area power grid.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Distributed lqr design for identical dynamically coupled systems: Application to load frequency control of multi-area power grid

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.230310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.194448Z digest=sha256:ae7934f313cf2f22f53a13cc1a1aa2d43818eba0d232fca3a74050957de31371

Observation d9516177-5a22-4215-8422-76abd2bc77a3 · outbound

This paper cites First-order regret in reinforcement learning with linear function approximation: A robust estimation approach.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon First-order regret in reinforcement learning with linear function approximation: A robust estimation approach

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.081763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.316145Z digest=sha256:bfc46db27b4373387eb7d2fe107750bc73825987d7f97ef4470d057f191cec04

Observation 0771b835-b586-4bf6-bf9a-a262f3fbfb0f · outbound

This paper cites Residual Bootstrap Exploration for Bandit Algorithms.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Residual Bootstrap Exploration for Bandit Algorithms

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:35:27.961339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.400756Z digest=sha256:d3209999e5282fb3db929305116389d272e6c0abfab2f358eae36ba7d043a3b7

Observation d0e2f6e8-73f4-4fb0-a457-ff4d79088e7b · outbound

This paper cites Exact asymptotics for linear quadratic adaptive control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Exact asymptotics for linear quadratic adaptive control

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.929421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.481427Z digest=sha256:8f2dfcdc8e0821ed627ce621be7c473e166f86e8adaebed19e97d5f11a4dd9be

Observation 8503830b-2585-4327-9b3c-0e64e498b5cf · outbound

This paper cites Continuous-time mean--variance portfolio selection: A reinforcement learning framework.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Continuous-time mean--variance portfolio selection: A reinforcement learning framework

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.818138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.551758Z digest=sha256:0ce65c232504b426291d46a241bbfcf34917b344968cc5688200627bb94d32d7

Observation 2331a859-145c-48a3-9385-e20128b4e3ea · outbound

This paper cites Stochastic zeroth-order optimization in high dimensions.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Stochastic zeroth-order optimization in high dimensions

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.720661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.632695Z digest=sha256:c3ddd3edf14451cc00958a77daf53e633f9dc99c55c1085d9b054b6dc24f9af3

Observation adc91b51-daa4-48e6-b62d-13db23fe28cf · outbound

This paper cites Leveraging linear quadratic regulator cost and energy consumption for ultrareliable and low-latency iot control systems.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Leveraging linear quadratic regulator cost and energy consumption for ultrareliable and low-latency iot control systems

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.601244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.691248Z digest=sha256:0a13ea299adf834fdc3974c06e7e6a08daf1b7105659658bf46aafbe1be8af8f

Observation e1e7ea49-2d11-44bf-b1b8-8a2588cd0fe3 · outbound

This paper cites Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.461702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.726098Z digest=sha256:65205d9b5d4bc31d823ac3de472d4b9f34fb9966fe73845d7a4c58fd64fe8460

Observation 1456cc1b-6615-44cd-b6f3-899786ffcef2 · outbound

This paper cites Online covariance matrix estimation in stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online covariance matrix estimation in stochastic gradient descent

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.305913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.778578Z digest=sha256:388a3ffbb16fb1ebc53b24e9da058941ba3d927eec0d561fdb55eceb0b49b0a3

Observation a418e2f7-6b01-417e-86f2-07c3a8339bae · outbound

This paper cites Uncertainty quantification and exploration for reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Uncertainty quantification and exploration for reinforcement learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.187289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.824202Z digest=sha256:3a4c342d044f7a5f5a4292a8e8fa1941ea68ff3b77c5a29a4ed621154dce8d91

Pith citing papers

No inbound Pith citation observations are available.