Pith. sign in

Paper Citation Record · LEDGER

Projection-Based Constrained Policy Optimization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2010.03152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.03152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:20:29.464007Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.827587Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c45262ad-fe4e-41fd-b2cf-23cf9243c6ec · inbound

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning cites this paper.

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:29.464007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:29.464007Z digest=sha256:39910f2630ecd17daeb1822fb8bf1d4b864075f2d27334fd8207be14a3c16806

Observation 69baed02-188c-4f37-84b3-29394e97d553 · inbound

Proactive Constrained Policy Optimization with Preemptive Penalty cites this paper.

Proactive Constrained Policy Optimization with Preemptive Penalty Projection-Based Constrained Policy Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:39.944282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:39.944282Z digest=sha256:33a8d0b238872b0ffa553d51c31a8405ee2dd91d7abd5e51cc5e7ef98b404ff3

Observation 5bd81fc8-9bc4-4a4e-ba30-9110a2ae3d83 · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework Projection-Based Constrained Policy Optimization

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:42.568008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:42.568008Z digest=sha256:20182d9f452121947445d1aa95c18de4fb432246b891f9cbd91e4556c58146b7

Observation 54f2db2d-3e3b-4647-bc16-4aba70695daa · inbound

HAEPO: History-Aggregated Exploratory Policy Optimization cites this paper.

HAEPO: History-Aggregated Exploratory Policy Optimization Projection-Based Constrained Policy Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:08.329208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:08.329208Z digest=sha256:2cf1bf329d290606f382c2c671a7b23d191501784044fdb87ae1e8957258b6cc

Observation 9d66733f-0c55-4b67-b2d9-06229e753e3c · inbound

Learning Reachability of Energy Storage Arbitrage cites this paper.

Learning Reachability of Energy Storage Arbitrage Projection-Based Constrained Policy Optimization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:11:23.314646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:09:57.342909Z digest=sha256:8d7e4d009e3fbf0596c160a003d5aece9c5d6d813f3a54468fc73c1ea923030d

Observation 74cbb9dd-b34a-4e9a-a670-d1f326ae747b · inbound

Shaping Zero-Shot Coordination via State Blocking cites this paper.

Shaping Zero-Shot Coordination via State Blocking Projection-Based Constrained Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.202198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:42:23.260464Z digest=sha256:6acafa2f85da0af25faec4a3ed31c8252fd82feda2ab9939cefec482dbcbdbc7

Observation 3707cb9f-bf1c-4307-94ab-185a1bb1573f · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.563478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:28:24.455817Z digest=sha256:91deae2023146d3dbd8fed03ed712932c32951cefe9d0fd814947a8395db8b0b

Observation 377cdaa2-8b56-4df6-8339-484b1c0cc31d · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.091288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T22:48:55.661356Z digest=sha256:a91253ef00788fdea43209c2e0c8da7e2631c916659517758de0975552b56b66

Observation b65ad3bb-b558-49a0-a9b5-dbc898cded5a · inbound

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation cites this paper.

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation Projection-Based Constrained Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.697736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T04:52:48.491036Z digest=sha256:a522c347bf0e1cf29204ccea4920ca49a2a8792828a15898f006fe116544d7b2

Observation 5302c802-cf66-4d30-9860-41e11ba0c694 · inbound

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability cites this paper.

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability Projection-Based Constrained Policy Optimization

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.836749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:50:40.285132Z digest=sha256:900bc58c251692c35b63145cc7b07f689021684173aa5e973e5ac5c4650c9410

Observation f72318d1-fc15-43de-9c53-e04372183f8b · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Projection-Based Constrained Policy Optimization

Reference 193

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.828921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:579d2ace3757b62f5198b579e317699f277f3649e66f0a7de9e18735b0696633

Observation 72aa2f1c-96ae-495b-a307-7c421392e84d · inbound

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control cites this paper.

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control Projection-Based Constrained Policy Optimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.387161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:49:22.328169Z digest=sha256:82fa7ec2d7ccbf98dc50470a46647340c8a47aab50bcd273288753a7e06b8fc9

Observation 358b8717-05b9-4301-89fc-6bd17154ec29 · inbound

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives cites this paper.

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives Projection-Based Constrained Policy Optimization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.013454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:27:10.983662Z digest=sha256:1c8ad169667654f1330244f289b5a36ee54f45ce24ac20a251a0de9750bf5280