Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:1810.07900.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:21:29.414918Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T04:16:36.072471Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ad02c292-1ac6-4c13-b308-fce82cf1bde8 · inbound
Optimistic Proximal Policy Optimization Policy Gradient in Partially Observable Environments: Approximation and Convergence
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eeb51ead-42fa-4bba-9a05-692da6fbf810 · inbound
Action-Gradient Monte Carlo Tree Search for Non-Parametric Continuous (PO)MDPs Policy Gradient in Partially Observable Environments: Approximation and Convergence
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e3dabbf-8f89-4c53-80c7-b30bff4f7e6c · inbound
Learning Deterministic Policies with Policy Gradients in Constrained Markov Decision Processes Policy Gradient in Partially Observable Environments: Approximation and Convergence
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f673fe8b-89e0-4bbd-8811-6c630ad2d443 · inbound
Evolutionary Optimization of Deep Learning Agents for Sparrow Mahjong Policy Gradient in Partially Observable Environments: Approximation and Convergence
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f294dcf-1119-4d71-b277-890e767a30c7 · inbound
The Value Function Semi-Algebraic Set in Partially Observable Markov Decision Processes Policy Gradient in Partially Observable Environments: Approximation and Convergence
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.