Pith. sign in

Paper Citation Record · LEDGER

Projection-Based Constrained Policy Optimization

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2010.03152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.03152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:20:29.464007Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.827587Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c45262ad-fe4e-41fd-b2cf-23cf9243c6ec · inbound

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning cites this paper.

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:29.464007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:29.464007Z digest=sha256:e78a6110af04cf366c702b071b19ed41b5224dbbe224b023c16325d482df4d6c

Observation 69baed02-188c-4f37-84b3-29394e97d553 · inbound

Proactive Constrained Policy Optimization with Preemptive Penalty cites this paper.

Proactive Constrained Policy Optimization with Preemptive Penalty Projection-Based Constrained Policy Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:39.944282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:39.944282Z digest=sha256:c03419ba6e124532f27723bed15af7f649d9b9ef504388e3f4dc7e68e889e2de

Observation 5bd81fc8-9bc4-4a4e-ba30-9110a2ae3d83 · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework Projection-Based Constrained Policy Optimization

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:42.568008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:42.568008Z digest=sha256:7ac1c59c28fa451f5efc0a07e9642071282f11004deab2a7b206b4a55f45148c

Observation 54f2db2d-3e3b-4647-bc16-4aba70695daa · inbound

HAEPO: History-Aggregated Exploratory Policy Optimization cites this paper.

HAEPO: History-Aggregated Exploratory Policy Optimization Projection-Based Constrained Policy Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:08.329208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:08.329208Z digest=sha256:505ad50dffd411af38e898cf56fcf73e3611b8b50dbd2347cb3b7cf4e9d05567

Observation 9d66733f-0c55-4b67-b2d9-06229e753e3c · inbound

Learning Reachability of Energy Storage Arbitrage cites this paper.

Learning Reachability of Energy Storage Arbitrage Projection-Based Constrained Policy Optimization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:11:23.314646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:09:57.342909Z digest=sha256:475f3b530a486b62e4a9a0db382a79612e9b1dcadff9e3643ccb46eda8175d14

Observation 74cbb9dd-b34a-4e9a-a670-d1f326ae747b · inbound

Shaping Zero-Shot Coordination via State Blocking cites this paper.

Shaping Zero-Shot Coordination via State Blocking Projection-Based Constrained Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.202198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T07:42:23.260464Z digest=sha256:141f176e828429bb642c27602bf340a1d1d0aa1a037772bdbb8af9c79a0cfd08

Observation 3707cb9f-bf1c-4307-94ab-185a1bb1573f · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.563478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T07:28:24.455817Z digest=sha256:6928e37884e7366de5d3eb945b7163e910f907d24acc9a7dccfe0f07c460cf87

Observation 377cdaa2-8b56-4df6-8339-484b1c0cc31d · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.091288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T22:48:55.661356Z digest=sha256:a4aa02c05191d6000e72d240a246f7845bbdf603dd0f484905e57990fe1957e6

Observation b65ad3bb-b558-49a0-a9b5-dbc898cded5a · inbound

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation cites this paper.

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation Projection-Based Constrained Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.697736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T04:52:48.491036Z digest=sha256:8508b8cdefd62e995aec6ef273884de1971e8c4a539e726e7d87fa71311c22aa

Observation 5302c802-cf66-4d30-9860-41e11ba0c694 · inbound

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability cites this paper.

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability Projection-Based Constrained Policy Optimization

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.836749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T02:50:40.285132Z digest=sha256:928e1b3dcbfd22f0ec78f0f15539e8c48316d3ec2ee3b911fc6800827d2395a4

Observation f72318d1-fc15-43de-9c53-e04372183f8b · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Projection-Based Constrained Policy Optimization

Reference 193

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.828921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:be1d61732796d13c6c52cf5da617283057268da64693ca8ea96931be73edbc86

Observation 72aa2f1c-96ae-495b-a307-7c421392e84d · inbound

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control cites this paper.

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control Projection-Based Constrained Policy Optimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.387161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T04:49:22.328169Z digest=sha256:577908fd5545976b8c94959ebe165b502ee5c3ce4a1ef84a3588aaa7a3c95961

Observation 358b8717-05b9-4301-89fc-6bd17154ec29 · inbound

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives cites this paper.

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives Projection-Based Constrained Policy Optimization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.013454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T04:27:10.983662Z digest=sha256:eab2b29cf16dd773e2af5f8b6abffddd3ad15025f442718bc428a07e33f88dbd