Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2005.03557.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:39:04.083982Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:49:42.782845Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c7fc9279-0d8b-412d-b67e-512d9eeca24c · inbound
Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 21ef55fd-364f-4bea-8ba7-d04d2b8f7bea · inbound
Decoupled Functional Central Limit Theorems for Two-Time-Scale Stochastic Approximation Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3378bad-de8b-4c30-b78e-c7e048cbb534 · inbound
Global Optimality of Single-Timescale Actor-Critic under Continuous State-Action Space: A Study on Linear Quadratic Regulator Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35f8e9cf-80a4-4bcb-9073-2f4fe6170e35 · inbound
Finite-Time Global Optimality Convergence in Deep Neural Actor-Critic Methods for Decentralized Multi-Agent Reinforcement Learning Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21564185-c876-47dc-8331-4d0d8d2f8743 · inbound
SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement Learning Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b993073-af06-439f-9818-87207d0e82cb · inbound
Optimal Sample Complexity for Single Time-Scale Actor-Critic with Momentum Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1008e2cc-2969-42a8-8d1e-1400f0862806 · inbound
Achieving $\epsilon^{-2}$ Sample Complexity for Single-Loop Actor-Critic under Minimal Assumptions Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ff79197d-a05b-4276-b55e-211b91379f23 · inbound
Stationary Robust Mean-Field Games under Model Mismatches Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f95e448a-d11d-43dd-852d-e375044c556e · inbound
On the Policy Convergence of Policy Mirror Descent Methods Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d88536b2-f16f-4da6-955f-bbe6e9d0198e · inbound
Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.