Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:1906.01786.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:41:31.745985Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:49:42.705118Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 20459a90-169e-428f-8f2e-858c58f9415f · inbound
Neural Policy Gradient Methods: Global Optimality and Rates of Convergence Global Optimality Guarantees For Policy Gradient Methods
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e97a6a-4784-4ae1-a00a-61754fd6ac25 · inbound
Adaptive Trust Region Policy Optimization: Global Convergence and Faster Rates for Regularized MDPs Global Optimality Guarantees For Policy Gradient Methods
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e89928d0-be98-41fd-bd19-ef0392e64aaa · inbound
Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Global Optimality Guarantees For Policy Gradient Methods
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05fc2478-8a1a-49cc-8652-54b738a9b238 · inbound
Imitate Optimal Policy: Prevail and Induce Action Collapse in Policy Gradient Global Optimality Guarantees For Policy Gradient Methods
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0935bf59-cf1b-4e87-94fa-078009b148f4 · inbound
Stationary Robust Mean-Field Games under Model Mismatches Global Optimality Guarantees For Policy Gradient Methods
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.