Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:07:09.566879Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 2 inbound Pith citation observations for arXiv:2506.00700.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:07:09.566879Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:37:41.630698Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-10T11:35:19.078046Z
14 of 14 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 074502be-9dda-46de-81e9-5d3b8d32d5fd · outbound
Central Path Proximal Policy Optimization A safe exploration approach to constrained Markov decision processes
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78af1034-67c6-4a9f-8fd3-24376f09bebb · outbound
Central Path Proximal Policy Optimization Direct Behavior Specification via Constrained Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4faecdac-4fa9-411c-a493-11ed459e1bb1 · outbound
Central Path Proximal Policy Optimization Responsive Safety in Reinforcement Learning by PID Lagrangian Methods
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d55bf51f-2cc4-4f0f-8774-b0fda9159b60 · outbound
Central Path Proximal Policy Optimization Constrained Reinforcement Learning with Smoothed Log Barrier Function
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3628220-eedf-49dc-a3a3-71f7824ecd61 · outbound
Central Path Proximal Policy Optimization Yiming Zhang, Quan Vuong, and Keith Ross
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d638274-83d0-44c6-8aef-e4d9f5cc284b · outbound
Central Path Proximal Policy Optimization surrogate advantage trick
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da807287-27f3-458c-99d7-ddad26648efb · outbound
Central Path Proximal Policy Optimization High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1d34400-fb63-46a9-b893-1e669deacd40 · outbound
Central Path Proximal Policy Optimization Dimitri P Bertsekas
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f78ce86-e4d7-4720-8ad6-c368ea7bfc38 · outbound
Central Path Proximal Policy Optimization Lyapunov-based Safe Policy Optimization for Continuous Control
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5a27ad-336f-436e-9782-1670098eb2d9 · outbound
Central Path Proximal Policy Optimization Jincheng Mei, Chenjun Xiao, Bo Dai, Lihong Li, Csaba Szepesvári, and Dale Schuurmans
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba153c4-cc06-4dc8-bf5c-d9d8f999b133 · outbound
Central Path Proximal Policy Optimization Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b21ca17-83c7-407c-bef8-f3c3fc4823e1 · outbound
Central Path Proximal Policy Optimization A unified view of entropy-regularized Markov decision processes
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eb6c714-6dfd-4a21-b04b-ee6f1af7c67d · outbound
Central Path Proximal Policy Optimization On PI Controllers for Updating Lagrange Multipliers in Constrained Optimization
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc27e03e-c0c5-4d02-a012-6a3d05bf8902 · outbound
Central Path Proximal Policy Optimization Embedding Safety into RL: A New Take on Trust Region Methods
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a1ef6a6-4859-4cd6-84de-2a19d6439ba4 · inbound
The Geometry of Nonlinear Reinforcement Learning Central Path Proximal Policy Optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d095a95e-6a3c-48b7-817c-c5acf95f3998 · inbound
Bounded Ratio Reinforcement Learning Central Path Proximal Policy Optimization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.