Pith. sign in

Paper Citation Record · LEDGER

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator

As of 16 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 1 inbound Pith citation observation for arXiv:2505.01041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01041 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:39:04.102807Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:19:12.496946Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:16:12.256474Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact1
  • verified fuzzy27
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 77beed5a-3374-4197-b449-10b64429b004 · outbound

This paper cites Optimal control: linear quadratic methods.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Optimal control: linear quadratic methods

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.198954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.933597Z digest=sha256:27110c5072fcec8d7077d8370158511e391796fc23ec7899ee3c1964c271881f

Observation 95e5b361-ca27-40dc-ad64-77c64c1f73d7 · outbound

This paper cites Analysis of a target-based actor-critic algorithm with linear function approximation.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Analysis of a target-based actor-critic algorithm with linear function approximation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.183451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.938760Z digest=sha256:40998b0cc8d853f91352670a6fa7241b74d58393181f7c0d2a9ffa8de4605968

Observation 1285b3ce-bc7e-4e0b-b7a1-6cc8aee2e572 · outbound

This paper cites Natural actor--critic algorithms.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Natural actor--critic algorithms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.169532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.943471Z digest=sha256:5c6e516421afaf8812e99206a1e9a559b00c8944a1c98f9759d79568d3db7a9d

Observation 07e00a86-c27a-4d1b-8eb8-d2e70870f18b · outbound

This paper cites Adaptive linear quadratic control using policy iteration.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Adaptive linear quadratic control using policy iteration

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.154817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.949259Z digest=sha256:84b7d1ef40b62e729c3189675ddda3150693a543f649f3eca81d8965d0559701

Observation dca7fea3-6cfd-43ff-998c-54b5142d5f6e · outbound

This paper cites A convergent online single time scale actor critic algorithm.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator A convergent online single time scale actor critic algorithm

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.141308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.953848Z digest=sha256:d7bd8bdd732f659dcfe6227d07c5545e3bbae08f610edd3456c9228fd7ea18f8

Observation 732c5bc3-1c8d-4570-b12f-8d0882aa34d2 · outbound

This paper cites Finite-time analysis of single-timescale actor-critic.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Finite-time analysis of single-timescale actor-critic

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:39:04.189631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.958608Z digest=sha256:7c78df4cd47c6503fb8f95015d9d1c11fea91273ac12cfa9a925ac999d823f33

Observation 111a90a9-f69c-4a9a-92bc-d171ff0dce6c · outbound

This paper cites Closing the gap: Tighter analysis of alternating stochastic gradient methods for bilevel problems.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Closing the gap: Tighter analysis of alternating stochastic gradient methods for bilevel problems

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.127660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.963866Z digest=sha256:303c702c1deac9f0765d32e4f61ac2541ce91bfbd083700e91f2fdde43be8210

Observation 3f906a92-ba3f-4522-88c2-d24a7aef8e0c · outbound

This paper cites Learning linear-quadratic regulators efficiently with only T regret.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Learning linear-quadratic regulators efficiently with only T regret

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.114199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.968112Z digest=sha256:c2daab7252c47691a814e3883c4b53b6bbc9d85e61171fefd6d65847275ccb0a

Observation 20e8ab97-7c29-4169-956a-c6af3c57576a · outbound

This paper cites Regret bounds for robust adaptive control of the linear quadratic regulator.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Regret bounds for robust adaptive control of the linear quadratic regulator

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:05.075198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.972346Z digest=sha256:baf19c5b73c3037f2cc1704618b9a31d40dd35195e20421d45e523597e7d07ff

Observation ecb9ae34-d866-4c73-936e-3d4fdb2b8883 · outbound

This paper cites On the sample complexity of the linear quadratic regulator.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator On the sample complexity of the linear quadratic regulator

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:03.976504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:03.976504Z digest=sha256:f361b5a59d1daa95f238fcb5f474a907dbb24afb5ddc7e34d4c6bae6a15368bf

Observation aa3a1011-1731-4097-ba84-07440835bf8e · outbound

This paper cites Optimization landscape of policy gradient methods for discrete-time static output feedback.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Optimization landscape of policy gradient methods for discrete-time static output feedback

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.932099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.980785Z digest=sha256:d2a3535805183a10a9221a22a5bae48d1cc7e06b3c206b1ef21bfdd7da069963

Observation 2e84f3da-7edf-4855-a6b8-9ebeb84828c3 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Global convergence of policy gradient methods for the linear quadratic regulator

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.907386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.984934Z digest=sha256:8f060a240368b1b2e27f94c4486500a00582a69044ab4a5c84efa9ef2f307e49

Observation 677c7683-b0de-40c4-9f58-45fd1a733706 · outbound

This paper cites On statistical characteristics of the product of two correlated chi-square variables.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator On statistical characteristics of the product of two correlated chi-square variables

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.892862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.989074Z digest=sha256:83ae94b4ea4e466e5b0d8c7823bc482c91bb8c9f53fcbec889853316e30b2f28

Observation 401322f9-ffdf-4bed-b316-cc952c734547 · outbound

This paper cites A natural policy gradient.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator A natural policy gradient

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.878879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.992826Z digest=sha256:d8565fd62b9026db36b12438b8272812f944d6b78595e6bd668e8a3f402affaa

Observation 36791a7f-7509-4adb-af9c-7931550b7683 · outbound

This paper cites Two time-scale stochastic approximation with controlled markov noise and off-policy temporal-difference learning.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Two time-scale stochastic approximation with controlled markov noise and off-policy temporal-difference learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.863882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:03.996901Z digest=sha256:ede4d6fcec333362200a913de53d2151f5a2fc29168bfaacb34c4cb7292cadd5

Observation b505173e-1a01-46e8-86a4-4125755e5b40 · outbound

This paper cites Actor-critic algorithms.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Actor-critic algorithms

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.850485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.000998Z digest=sha256:d6d2e82c2543358827a756aedeebd5df57e3cce675ef7cf5d3a64f220190e9b4

Observation f32f9f7d-1bab-4bf0-877c-4850136440df · outbound

This paper cites Finite-time analysis of approximate policy iteration for the linear quadratic regulator.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Finite-time analysis of approximate policy iteration for the linear quadratic regulator

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.832550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.005015Z digest=sha256:589b1bc137c3f618d6956cee2889ff469d50bf353db3feb817ce1d1017da8942

Observation 76331952-d21d-4999-9597-0720a0dcf5e4 · outbound

This paper cites On the Sample Complexity of Actor-Critic Method for Reinforcement Learning with Function Approximation.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator On the Sample Complexity of Actor-Critic Method for Reinforcement Learning with Function Approximation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.009195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.009195Z digest=sha256:709d5778c48f152633402a7cc737a93166731e73aed3dbc20196dbcd9f31e985

Observation ef4b0f6b-9f73-4d8c-aab5-0cdaf956ae03 · outbound

This paper cites Deep learning.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Deep learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.014586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.014586Z digest=sha256:fefd8cbf68fc9d3eacccb76897c243049cc3461b932d1b1dc4f4159a2df4b472

Observation bcdf1311-62cd-4a86-a549-2b10df58d446 · outbound

This paper cites The moments of products of quadratic forms in normal variables.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator The moments of products of quadratic forms in normal variables

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.734062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.018899Z digest=sha256:d17d84f9494fc50a6c42d164757a978e76505a1e15b8fae16b7610acced779e5

Observation bb4c4a02-5450-48f6-884d-0ffaf396339a · outbound

This paper cites Derivative-free methods for policy optimization: Guarantees for linear quadratic systems.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Derivative-free methods for policy optimization: Guarantees for linear quadratic systems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.550748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.022973Z digest=sha256:8724e95714eed7f6d8a4b8c0479a0a52fb93598969a802d9528fc903b96664a2

Observation 35d03659-b21b-4470-970a-8db91b6d649b · outbound

This paper cites Certainty equivalence is efficient for linear quadratic control.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Certainty equivalence is efficient for linear quadratic control

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.027127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.027127Z digest=sha256:5b6f8f933cd3822fe260d7783bd96035fdf1728bc34367d51ec0972b8641afa4

Observation 01ca1eec-99a7-404e-b094-d9c6b172bdb3 · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Asynchronous methods for deep reinforcement learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.031146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.031146Z digest=sha256:d3bbb4131c128a9274fd7d6e87e62259fab36b3bab9863050f42c710b3e4a4a6

Observation 45617141-dd60-40e1-92e9-66064391c51e · outbound

This paper cites The bias and moment matrix of the general k-class estimators of the parameters in simultaneous equations.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator The bias and moment matrix of the general k-class estimators of the parameters in simultaneous equations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.361828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.035398Z digest=sha256:593a95e59279138028871d2a2ac1f52b8206769c99916eb9273ee80838035976

Observation 6451decf-8e6d-40e0-a40a-76eb2906cdb9 · outbound

This paper cites A small gain analysis of single timescale actor critic.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator A small gain analysis of single timescale actor critic

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.348797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.039811Z digest=sha256:c81e33d28b07d999746301035787d767cbf1f2481ca4bf2ab6441281f824eb88

Observation 375b99d3-a857-4d3e-a822-3f03f743f72c · outbound

This paper cites Linear models in statistics.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Linear models in statistics

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.335451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.043884Z digest=sha256:284c17db125bee89b658a22479b6f5934f0267b1c7fdccbff1248b40b7201126

Observation df3c4068-ccd1-4e46-b7c6-dcba03e0aab2 · outbound

This paper cites On the kronecker product.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator On the kronecker product

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.322080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.048063Z digest=sha256:94a45ac86cc14d24ff0fe4291b34907b6e664f8305673b91edc6e08b3ea6d638

Observation 478cbc88-863d-44f8-9525-424162a64ad9 · outbound

This paper cites Trust region policy optimization.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Trust region policy optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.052431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.052431Z digest=sha256:ed1956e759b5584f1beaf93e297e8f6cf82bf6c8d649e22262f7a0147566fedd

Observation a4f946c3-44e1-4c6d-818e-f5fe2099ffd1 · outbound

This paper cites Mastering the game of go without human knowledge.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Mastering the game of go without human knowledge

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.300132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.056667Z digest=sha256:b78a7f64c3c925fbd6506f7a233b1314d4b0c9e8abf1bdf0afe6e1043121f528

Observation feeaa4a2-c026-49df-b038-72b8ab5fb6a8 · outbound

This paper cites Reinforcement learning: An introduction.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Reinforcement learning: An introduction

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.060747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.060747Z digest=sha256:18606f952958a72aaa57fb6bc84e24b0cdf4c88160448c78f4668a7726d12a7c

Observation 5573521a-383a-4246-a56c-00624d160314 · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Policy gradient methods for reinforcement learning with function approximation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.066728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.066728Z digest=sha256:2fd5043a020351e69d859ca18afcbaa87dcc2a5f95790c02254778d7c8196ffa

Observation 1d99fc50-763c-44f9-a6b6-7c82327147c5 · outbound

This paper cites Least-squares temporal difference learning for the linear quadratic regulator.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Least-squares temporal difference learning for the linear quadratic regulator

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.269001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.071051Z digest=sha256:9af501ab02aa929d06ea925d35ed690fe08491cb51cf20891da2a23c867bbd56

Observation e65fee3c-47ea-4add-b470-0f22216c3445 · outbound

This paper cites Neural Policy Gradient Methods: Global Optimality and Rates of Convergence.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Neural Policy Gradient Methods: Global Optimality and Rates of Convergence

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.075074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.075074Z digest=sha256:0a19dfc1e34c4fcfb854145a78364eaab01dff4550e9672f51655e7f7334bb8a

Observation 1d50a4fe-ed6d-490a-a5f6-920ac378f22a · outbound

This paper cites A finite-time analysis of two time-scale actor-critic methods.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator A finite-time analysis of two time-scale actor-critic methods

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.255580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.079544Z digest=sha256:21a1f49418dbccfbdfbd54d8802f6984808c55e88752f0e72d424a9ef9b1cc59

Observation a3378bad-de8b-4c30-b78e-c7e048cbb534 · outbound

This paper cites Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.083982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.083982Z digest=sha256:ceece817fa40a72847b0cd77e4cbfec1475f339de3d9126537038eced4ba08bb

Observation 7b3dcbe4-54e8-475b-9c6c-b6937a689ed9 · outbound

This paper cites Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.242143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.089531Z digest=sha256:4ca5e9be09865a110b1dd22d5e4deb90bd95302753b1674f54f378a088815320

Observation bdcce64d-f4f4-4b70-a10c-6f7fd701b328 · outbound

This paper cites Provably convergent two-timescale off-policy actor-critic with function approximation.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Provably convergent two-timescale off-policy actor-critic with function approximation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.227283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.093831Z digest=sha256:d902ad32558407a7a6a8b76d64e1e066a4e7bcfd27a697e1535bf0c1ed859e43

Observation 1bdb2454-131d-4cbc-9021-1ee40c1bb6fd · outbound

This paper cites Single timescale actor-critic method to solve the linear quadratic regulator with convergence guarantees.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Single timescale actor-critic method to solve the linear quadratic regulator with convergence guarantees

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:39:04.213778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T04:39:04.098363Z digest=sha256:3ce7917445064a050fcaae119d94a73cac99fb4cc17e7286d1c692c6177e0516

Observation 33689e7b-af1a-481d-843f-7dee1f23cbb5 · outbound

This paper cites write newline.

Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator write newline

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T04:39:04.102807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:39:04.102807Z digest=sha256:69ef35cce48be4f85b669b6f59492f5bb70762b9c331efbc74ae697aeca9a396

Pith citing papers

Observation f137a911-33be-4446-b838-e91cae6a39b3 · inbound

Model-free LQG Control with Chance Constraints cites this paper.

Model-free LQG Control with Chance Constraints Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.258406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T21:19:12.496946Z digest=sha256:41ec038cdd3c85c5fc654f3847c995cf54f6376b1e02711bf71fb5f7797f8a4e