Pith. sign in

Paper Citation Record · LEDGER

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon

As of 20 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2508.08436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.08436 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:35:27.824202Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy47
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60bb9e6b-52c1-4116-9982-3b30e9f26cf4 · outbound

This paper cites Normal approximation for stochastic gradient descent via non-asymptotic rates of martingale clt.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Normal approximation for stochastic gradient descent via non-asymptotic rates of martingale clt

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.674006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.798849Z digest=sha256:fc2f79e99198a62269ec05cf5d212f8298259d13eeaca69ea0dcfefd97b84d51

Observation 3199ac5b-498f-4ef3-9568-7f0e4bdfbafb · outbound

This paper cites Adaptive control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.464013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.844321Z digest=sha256:3c3f172946d434e09850e8e2caef5623bb0c7429944db7ebcfeb307077d09729

Observation ba82c5a5-668d-4e0b-9964-43deed480fbc · outbound

This paper cites Logarithmic regret for episodic continuous-time linear-quadratic reinforcement learning over a finite-time horizon.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Logarithmic regret for episodic continuous-time linear-quadratic reinforcement learning over a finite-time horizon

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.327979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.885776Z digest=sha256:57b7bc5fc606488334549eed805eea1add9e5cd18fbac400d778fa712f1e221c

Observation d460d6e4-d5e4-47c8-bee7-cdd8a10eb7c0 · outbound

This paper cites Adaptive control with the stochastic approximation algorithm: Geometry and convergence.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Adaptive control with the stochastic approximation algorithm: Geometry and convergence

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:36.121599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:23.974853Z digest=sha256:765fcbc91c1fa2b240d1a86da64a1aee3137aefd2e5753223d7892d75139500a

Observation aaad1fa6-b5a3-43de-927a-534f9dcfc526 · outbound

This paper cites Dynamic programming and optimal control: Volume I , volume 4.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic programming and optimal control: Volume I , volume 4

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.925538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.061034Z digest=sha256:4c218eb73e7c97771b0f081dd60d2b358045b8dd89c5fde62d7fa18edbaec6c5

Observation 9b498ff2-d462-4347-bcde-18f845d587a2 · outbound

This paper cites Reinforcement learning applied to linear quadratic regulation.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning applied to linear quadratic regulation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.688872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.177822Z digest=sha256:9a092e86cec3b3e5d2e90f5b4238c4503d8c9d01a48dde06b26f0a7577aa1fcd

Observation 1157fc72-6af4-465d-b37b-4467c5d41ba4 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:24.252521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:24.252521Z digest=sha256:66d404900b9df00b163d90b31c92797054e1ad1602e54642200d12578e771a45

Observation 9542fdcd-b05b-4aa5-b292-357ad7a7166d · outbound

This paper cites Statistical inference for online decision making via stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for online decision making via stochastic gradient descent

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.529353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.312545Z digest=sha256:34e8df516df956415358135ebf3d274f5648b394dcb094099d81ba76702de793

Observation 39472f0f-c982-4e40-b087-97cc71a7fa31 · outbound

This paper cites Statistical inference for model parameters in stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical inference for model parameters in stochastic gradient descent

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.356129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.351893Z digest=sha256:a58ffb0d142bcf386bd6b02109cff373488a9ec0e3cd9992a43e9cff3ea92606

Observation c2c11bd4-9725-4966-ba3b-3c76681f879b · outbound

This paper cites Robust inference via multiplier bootstrap.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust inference via multiplier bootstrap

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:35.130232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.431215Z digest=sha256:e0b96f45860de88966c2688a7755b8f2f477f7b401ee026eb6bbc8900c048641

Observation 618178cc-5fd5-4d0c-90fc-dc6a9eebd61d · outbound

This paper cites On the sample complexity of the linear quadratic regulator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the sample complexity of the linear quadratic regulator

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:24.507547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:24.507547Z digest=sha256:eadc5b3be137e84b1bb3356d96eb1b7aa6336499c504217291b816adac277a35

Observation a2d1892e-a99d-49fe-a8bd-3a3e62f4c9fe · outbound

This paper cites Challenges of reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.966921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.579776Z digest=sha256:1b769e0546b890f951bac003cbda120a67cc420500aac84e134d1f0d4926ee86

Observation 61a1fb16-e1db-4c3e-a0b2-89e0f6070336 · outbound

This paper cites Challenges of real-world reinforcement learning: definitions, benchmarks and analysis.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Challenges of real-world reinforcement learning: definitions, benchmarks and analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.812742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.650888Z digest=sha256:a1b8050fd15874b2d2da192c01dd1d2d02be0c41f3d7f7356f0116c7c807da6b

Observation 16510445-8a5f-40b8-98ab-4543b24c3fb7 · outbound

This paper cites On the convergence theory of debiased model-agnostic meta-reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On the convergence theory of debiased model-agnostic meta-reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.551613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.720287Z digest=sha256:48eaf5db1eed8c64838452b8b4c388f72974230db3d6a5401f0343e38c962a49

Observation 5d085cb1-e113-402c-9a05-f1995d8cebc5 · outbound

This paper cites Online bootstrap confidence intervals for the stochastic gradient descent estimator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap confidence intervals for the stochastic gradient descent estimator

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.253477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.763782Z digest=sha256:6848ed9c210410ffc245936c9d5a128e7e3e5c8fc5b06e20aa3288ba24ae31fd

Observation 43d704f2-298b-4d3d-a850-d1d5cad672cc · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Global convergence of policy gradient methods for the linear quadratic regulator

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:34.063605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.816293Z digest=sha256:32fe3f48381cc7924489997897c5e58d358669403749ad0bd496924583e9ace8

Observation 784d7f5b-b5ee-4f7f-9fc0-2518b2c347cc · outbound

This paper cites Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.832945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.894628Z digest=sha256:3583595c2e8ee93dc763d4831e88c7ae4625be4c5c2b100508f1bf9b52179aec

Observation 56470e8b-ced4-404f-8470-f61f57a7ba36 · outbound

This paper cites Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.518251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:24.975470Z digest=sha256:33bcd8f76282630d0127b200a3f776cc58b43051d436b5495d40fbc9c117859b

Observation 5e84d419-1be6-43ef-a171-90223ace1d2d · outbound

This paper cites Policy gradient methods for the noisy linear quadratic regulator over a finite horizon.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Policy gradient methods for the noisy linear quadratic regulator over a finite horizon

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.307106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.083912Z digest=sha256:4255014d86373b02650c2b93d0ca8df1d3dd33636452f7fa5a3df1beaec42e4c

Observation 230fdaaf-92fe-43e4-b115-fa6a57d1d6fb · outbound

This paper cites Bootstrapping upper confidence bound.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Bootstrapping upper confidence bound

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:33.045307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.186107Z digest=sha256:9f0bebdeceb92802291c76fd40aca98d58fb7e83cf0a5cd8085912be7bb9b950

Observation 1824c0d6-9934-4a8b-875a-ce502789a0ba · outbound

This paper cites Model-based or model-free, a review of approaches in reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Model-based or model-free, a review of approaches in reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.809383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.271012Z digest=sha256:1841aab803935d717e39798113bf56a4280fa9c05d8e3e7acf8afe4b613a3c14

Observation d3bd7a1f-f21a-4334-be3d-23566858cd6c · outbound

This paper cites Improved zeroth-order variance reduced algorithms and analysis for nonconvex optimization.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Improved zeroth-order variance reduced algorithms and analysis for nonconvex optimization

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.496520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.382262Z digest=sha256:a659ddffcd3e8fc573780e9295cec43c2656a33a679f12a1d6bf1c3ebd02399e

Observation 66e8339b-58a2-4b8c-ba4f-0dbee8facd9e · outbound

This paper cites On incomplete learning and certainty-equivalence control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon On incomplete learning and certainty-equivalence control

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.213608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.472692Z digest=sha256:b9977d9f6674e47375552d9cadd739a4fa1a4c5b0b4921b99f31589273085d71

Observation bc17931f-76a1-4e95-8317-0c21a9ca6694 · outbound

This paper cites Iterated least squares in multiperiod control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Iterated least squares in multiperiod control

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:32.019692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.550290Z digest=sha256:cbb6f2f7ab9c0827ed8a8289aa1d2259fba8e15b83dc80d51e334ed9917b3568

Observation 2257be21-7815-492a-a897-19265deda77c · outbound

This paper cites Fast inference for quantile regression with tens of millions of observations.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Fast inference for quantile regression with tens of millions of observations

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.868085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.619163Z digest=sha256:61a2df58f225833ed7789e034124dce2ae59b8284a7a6b98598708a2eb751d39

Observation 589ef146-8f4a-4082-a5fc-dee4d511edaf · outbound

This paper cites Statistical Estimation and Inference via Local SGD in Federated Learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Statistical Estimation and Inference via Local SGD in Federated Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:25.666238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:25.666238Z digest=sha256:452ae7b6ed7ac447e65a8d3370087ea0a825d10303c9696e34a2d8040b158fa0

Observation 3d1b0510-8ed4-43e9-87d0-08bd5e84a753 · outbound

This paper cites Unifying offline causal inference and online bandit learning for data driven decision.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Unifying offline causal inference and online bandit learning for data driven decision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.661794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.727112Z digest=sha256:6ec9ce923599314155af34f4d6cf7db6e5eb5ac082906fca15a77540c3263ee5

Observation 4d034d63-7a58-4444-8405-6dffb8f0d0c5 · outbound

This paper cites A generalized reinforcement-learning model: Convergence and applications.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A generalized reinforcement-learning model: Convergence and applications

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.373166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.776992Z digest=sha256:94905d3a7b762402e3543069cf06256c1b65d6b6ae5a01b071cfea692fa847e1

Observation 97fa945a-a8a3-4591-bb2f-d8e8db676be7 · outbound

This paper cites A review of uncertainty for deep reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A review of uncertainty for deep reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:31.091172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.853537Z digest=sha256:654359c5328c3f0725abcfb9f1cab0a79a44adb61a05720a09a2a6f42a32f20e

Observation 3d0b37d9-9973-49f8-8287-d6ee78c67900 · outbound

This paper cites Gradient estimation in model-based reinforcement learning: a study on linear quadratic environments.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gradient estimation in model-based reinforcement learning: a study on linear quadratic environments

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.798966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:25.986287Z digest=sha256:5b88b47462786fb72c756eacdec1ec5146487fc3957ed549a4dd1c112529a392

Observation f34550ab-8f43-4611-bf02-7b9c047cc283 · outbound

This paper cites Certainty equivalence is efficient for linear quadratic control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Certainty equivalence is efficient for linear quadratic control

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.689389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.049926Z digest=sha256:74a2a688b39d1f1964e1d2ae2709fed4171acc103613dcfa717d2d389896720d

Observation 63d803c6-3821-42da-8ab8-8c9b79cc96e4 · outbound

This paper cites A linear quadratic regulator based speed control for remote-controlled racing cars.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon A linear quadratic regulator based speed control for remote-controlled racing cars

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.563763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.115250Z digest=sha256:b9ea72007f19789c94be61a945e0525e5b1914a4e2faf44394b1afc22cf1089b

Observation ee475c6d-d57e-4554-b11f-910186482ddd · outbound

This paper cites Black-box generalization: Stability of zeroth-order learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Black-box generalization: Stability of zeroth-order learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.405457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.193252Z digest=sha256:6121c241602fb34fef55e72c5580b059b0bb8fc39df41973cd028f49f055135b

Observation a09a9f0a-2c0a-41fd-8c4a-7349123df957 · outbound

This paper cites Acceleration of stochastic approximation by averaging.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Acceleration of stochastic approximation by averaging

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.283756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.241782Z digest=sha256:1914ff83c753345f50a56db9a3b7850db4401ec9ed3d7025da2f1cb473c8f47d

Observation 73f85c50-cd5d-4224-b957-f177ab39d3e0 · outbound

This paper cites Temporal Difference Models: Model-Free Deep RL for Model-Based Control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Temporal Difference Models: Model-Free Deep RL for Model-Based Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:26.295386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:26.295386Z digest=sha256:7881dc661ba72005ec18a038e5021ae0c5eec8cf6a852c887cd38da2cdb95928

Observation 70c83ace-2b1b-4688-a7de-a53ab20b67fd · outbound

This paper cites Online bootstrap inference for policy evaluation in reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online bootstrap inference for policy evaluation in reinforcement learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:30.143073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.387339Z digest=sha256:78fda46d165a272927407251073b03d772d175602441cdd68780a125a1fb8dc9

Observation 0a0be36e-29f1-4e77-8f9e-722e063a427e · outbound

This paper cites The unintended consequences of discount regularization: Improving regularization in certainty equivalence reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon The unintended consequences of discount regularization: Improving regularization in certainty equivalence reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.983187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.452152Z digest=sha256:d47c5cec862e0ec4efbd2a3fa5ffe437274ec686b2099d922c79bbfb73be66a9

Observation 33bd0d6e-8c77-488e-8e73-9b1ae9647dd4 · outbound

This paper cites Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Implicit Bias of Policy Gradient in Linear Quadratic Control: Extrapolation to Unseen Initial States

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:35:28.059162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.554802Z digest=sha256:83e6fbefcdbdad240ad214d1445b011b83c3068a3f18f3f47e9f185744245a9b

Observation 5b898f0e-802a-4fc5-a8c1-7309592af786 · outbound

This paper cites Efficient estimations from a slowly convergent robbins-monro process.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Efficient estimations from a slowly convergent robbins-monro process

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:26.641853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:26.641853Z digest=sha256:0ec67ed28a01a8a5dae6fa127877aee16a959f1a19c06173801478384982b82d

Observation 16618310-e5ab-42c0-a114-9808f6df6232 · outbound

This paper cites Gaussian approximation and multiplier bootstrap for polyak-ruppert averaged linear stochastic approximation with applications to td learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Gaussian approximation and multiplier bootstrap for polyak-ruppert averaged linear stochastic approximation with applications to td learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.854106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.704743Z digest=sha256:e1bad437b0854b16ff079b735859d8e4a98b1d433ac79153c3ec7499265925ef

Observation 9853206d-8b5a-4474-862e-f13eda4574c2 · outbound

This paper cites Dynamic causal effects evaluation in a/b testing with a reinforcement learning framework.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Dynamic causal effects evaluation in a/b testing with a reinforcement learning framework

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.720305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.808715Z digest=sha256:d5757e4a221e4d30ab9aa8602542e312973da72530459ac7deed633ccf343440

Observation f24ae1fc-815c-4c3c-a630-755deec6d106 · outbound

This paper cites Learning optimal controllers by policy gradient: Global optimality via convex parameterization.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Learning optimal controllers by policy gradient: Global optimality via convex parameterization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.598742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.899905Z digest=sha256:63650ce7811dbbf636306be1b333e268ec67f47138f101ad968e3d90e6d6f311

Observation 82b60fb8-3072-4679-a76d-90a8211c62c1 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Optimal scheduling of entropy regularizer for continuous-time linear-quadratic reinforcement learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.477974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:26.983368Z digest=sha256:9a706b2c81b7ace8bfd52b917f7224351fb22fbcc25d41b1300a0ccbc6d555bb

Observation ec2992d7-cfff-4afc-8363-59c56f6cce3f · outbound

This paper cites Robust exploration in linear quadratic reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Robust exploration in linear quadratic reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.368728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.100028Z digest=sha256:a87bde9e127deefd688844cfdcf562e831554402682137e0a7d66dafbf5b087e

Observation 1b34a2f0-c3bf-415c-bbaf-e3679dd4cd02 · outbound

This paper cites Distributed lqr design for identical dynamically coupled systems: Application to load frequency control of multi-area power grid.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Distributed lqr design for identical dynamically coupled systems: Application to load frequency control of multi-area power grid

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.230310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.194448Z digest=sha256:e8bc05c58a4153d2c84b101cee26eab31cd49e42d1291e4993ee33690efabe0e

Observation d9516177-5a22-4215-8422-76abd2bc77a3 · outbound

This paper cites First-order regret in reinforcement learning with linear function approximation: A robust estimation approach.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon First-order regret in reinforcement learning with linear function approximation: A robust estimation approach

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:29.081763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.316145Z digest=sha256:20b4dd66d3d8edeb9d75e17ea3155e513cbc1d026b314c7378adef799de84722

Observation 0771b835-b586-4bf6-bf9a-a262f3fbfb0f · outbound

This paper cites Residual Bootstrap Exploration for Bandit Algorithms.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Residual Bootstrap Exploration for Bandit Algorithms

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:35:27.961339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.400756Z digest=sha256:4dcc2dfb284da0241d2c584399d20a5d5ff9c2ce414a2cf31a416e08d064559e

Observation d0e2f6e8-73f4-4fb0-a457-ff4d79088e7b · outbound

This paper cites Exact asymptotics for linear quadratic adaptive control.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Exact asymptotics for linear quadratic adaptive control

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.929421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.481427Z digest=sha256:d981028ed8ddb57dda25b6644081b6a318ae982810e1de3213b6d08b667502ef

Observation 8503830b-2585-4327-9b3c-0e64e498b5cf · outbound

This paper cites Continuous-time mean--variance portfolio selection: A reinforcement learning framework.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Continuous-time mean--variance portfolio selection: A reinforcement learning framework

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.818138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.551758Z digest=sha256:87b60dfe99b254e1e459c297673ab64f11ab5636a460985d1e0b2ddbe132d6de

Observation 2331a859-145c-48a3-9385-e20128b4e3ea · outbound

This paper cites Stochastic zeroth-order optimization in high dimensions.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Stochastic zeroth-order optimization in high dimensions

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.720661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.632695Z digest=sha256:990cd72b6b11052e905a5d9d6182a43d2c9844a3fd61dce6b530c7cc36f2ee56

Observation adc91b51-daa4-48e6-b62d-13db23fe28cf · outbound

This paper cites Leveraging linear quadratic regulator cost and energy consumption for ultrareliable and low-latency iot control systems.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Leveraging linear quadratic regulator cost and energy consumption for ultrareliable and low-latency iot control systems

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.601244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.691248Z digest=sha256:811ac2bfd3cf87526ca308c50e7e676d31772911b370574b098af7aa4ba0b63a

Observation e1e7ea49-2d11-44bf-b1b8-8a2588cd0fe3 · outbound

This paper cites Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.461702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.726098Z digest=sha256:35453ea2465c04e990d9dda19b0a8f4f4754e6860db0985843bd9e03a7dc604d

Observation 1456cc1b-6615-44cd-b6f3-899786ffcef2 · outbound

This paper cites Online covariance matrix estimation in stochastic gradient descent.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Online covariance matrix estimation in stochastic gradient descent

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.305913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.778578Z digest=sha256:eb683469e807cc6cbd0d7fb4dce3d0de59155075a22b696ea9b50731117af1be

Observation a418e2f7-6b01-417e-86f2-07c3a8339bae · outbound

This paper cites Uncertainty quantification and exploration for reinforcement learning.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Uncertainty quantification and exploration for reinforcement learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:35:28.187289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T21:35:27.824202Z digest=sha256:b2cd1e757e46501e499ce8c2e8314d4b91ac4781c79b94a64adfda357fd5846f

Pith citing papers

No inbound Pith citation observations are available.