Pith. sign in

Paper Citation Record · LEDGER

Robust Peak-cost Constrained Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2607.15457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.15457 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T23:27:44.219249Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d550f721-d9aa-4d0b-bd27-d3900b92329d · outbound

This paper cites Altman,Constrained Markov decision processes.

Robust Peak-cost Constrained Reinforcement Learning Altman,Constrained Markov decision processes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:40.417193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:40.417193Z digest=sha256:42cdefc73d93f3677e16292a8a20a6b7a71573f3dde86d1c63a08e322b6bced2

Observation b61eda01-2514-4db4-9351-c36311815d6a · outbound

This paper cites Towards achieving sub-linear regret and hard constraint violation in model-free rl,.

Robust Peak-cost Constrained Reinforcement Learning Towards achieving sub-linear regret and hard constraint violation in model-free rl,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:40.496655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:40.496655Z digest=sha256:baa8446461d881359a99b3978fabcf71bad8b805ac37af0c50f49ee0a266ed40

Observation c26db9a6-4743-45bd-a319-814d4fa18fba · outbound

This paper cites Achieving sub-linear regret in infinite horizon average reward constrained mdp with linear function approximation,.

Robust Peak-cost Constrained Reinforcement Learning Achieving sub-linear regret in infinite horizon average reward constrained mdp with linear function approximation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:40.656413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:40.656413Z digest=sha256:4659f75fea88ab1905a585cee91a6dff21bc013c3f6498b38136d974789f4179

Observation e3d03652-b9b1-4f2b-9ddd-b72284d7251d · outbound

This paper cites Reward constrained policy optimization,.

Robust Peak-cost Constrained Reinforcement Learning Reward constrained policy optimization,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:40.743145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:40.743145Z digest=sha256:33c926eb8c9ed60b319da948cfe2a116e9ca13e2385e57ef78a58a03c5a7fc37

Observation 31754b4f-b1c8-4c3b-8456-2959d730a20f · outbound

This paper cites Reachability constrained rein- forcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Reachability constrained rein- forcement learning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:40.920680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:40.920680Z digest=sha256:f6d01b72c328f355c8b6c5179002747a27923831928618fd68ac7ef843734f8f

Observation a139e191-a795-4eba-936c-f7282e0165f7 · outbound

This paper cites Iterative reachability estimation for safe reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Iterative reachability estimation for safe reinforcement learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.058213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.058213Z digest=sha256:29a7bc22a71c8b2ead70c25d5e91e37cbabdd301fbda423de852fa4ca46a394d

Observation 38a08963-a5dd-46d6-ac79-c0d42d281c08 · outbound

This paper cites Safe policies for reinforcement learning via primal-dual methods,.

Robust Peak-cost Constrained Reinforcement Learning Safe policies for reinforcement learning via primal-dual methods,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.194510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.194510Z digest=sha256:31b96b0abb86f2190222a5c413e4fbc638b8d4fae776d1e3681accda51299907

Observation 37341d57-cd45-41c7-af60-ee570d854f3f · outbound

This paper cites Responsive safety in reinforce- ment learning by pid lagrangian methods,.

Robust Peak-cost Constrained Reinforcement Learning Responsive safety in reinforce- ment learning by pid lagrangian methods,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.347329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.347329Z digest=sha256:ad4ee975558182c4e5bbbfce4db62e6c53c66602dc1b3400b8bc0b08b7df5c72

Observation a9171b69-a913-40ce-8d5f-21f35d841a37 · outbound

This paper cites Constrained upper confidence reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Constrained upper confidence reinforcement learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.471554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.471554Z digest=sha256:bce11b7775d40a540dfa8a875a479f031924e4ea9eee6e29a23f430d221055c5

Observation d9c5dd50-a78f-436c-a92d-ef40ba3a59b5 · outbound

This paper cites Natural policy gradient primal-dual method for constrained markov decision pro- cesses,.

Robust Peak-cost Constrained Reinforcement Learning Natural policy gradient primal-dual method for constrained markov decision pro- cesses,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.636006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.636006Z digest=sha256:4a016822b6db84385b4f40ad5bd50f5d2822a370b6d39f105c68636bd875be17

Observation 5b4699e0-fdf6-4279-b728-10bc4cc72b61 · outbound

This paper cites Crpo: A new approach for safe reinforcement learning with convergence guarantee,.

Robust Peak-cost Constrained Reinforcement Learning Crpo: A new approach for safe reinforcement learning with convergence guarantee,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.754843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.754843Z digest=sha256:7a52387ec4be43d28df2bfbba0168aadc1694d392e9d126fd7dae9a2b1f4a0cb

Observation 5547b983-519f-4b5b-b9d4-c8ddff7baeeb · outbound

This paper cites Constrained policy optimization,.

Robust Peak-cost Constrained Reinforcement Learning Constrained policy optimization,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:41.899865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:41.899865Z digest=sha256:030111a1e8bb3558d242540ff1ec87d1d0553104c73ec8b881b36b8c73b88df3

Observation 0bb26c26-2813-4da9-a282-e92d4c3174cd · outbound

This paper cites A lyapunov-based approach to safe reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning A lyapunov-based approach to safe reinforcement learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.130824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.130824Z digest=sha256:bcdb71787d3c30bd8b0f869fe76c00bd6857f8519026d46c55d06be299b5134f

Observation 840bcb1f-cfb1-4b76-9202-9cb806b31959 · outbound

This paper cites Robust Constrained Reinforcement Learning.

Robust Peak-cost Constrained Reinforcement Learning Robust Constrained Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.254293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.254293Z digest=sha256:998e13dfdca19080b6d5a0e10029f1bb7c14fbee9065ce38fb956e61776592f6

Observation 14019e84-eacc-48bf-a23c-00787183dbe6 · outbound

This paper cites Distributionally Robust Constrained Reinforcement Learning under Strong Duality.

Robust Peak-cost Constrained Reinforcement Learning Distributionally Robust Constrained Reinforcement Learning under Strong Duality

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.392615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.392615Z digest=sha256:60218c9d095afe6f31562f3ee1ce155f0c6ec382b59d9fde66453889758fbb43

Observation ba752abe-9500-4c0a-9cb7-c48984a3d790 · outbound

This paper cites Near-optimal policy identification in robust constrained markov decision processes via epigraph form,.

Robust Peak-cost Constrained Reinforcement Learning Near-optimal policy identification in robust constrained markov decision processes via epigraph form,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.549293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.549293Z digest=sha256:69675c5d4c814b8343b9518ab871f2ef8d4faaf92d92203f4716fcedcc2a5435

Observation df4089d0-49e6-4644-890f-cd5e3e04693b · outbound

This paper cites Efficient policy optimization in robust constrained mdps with iteration com- plexity guarantees,.

Robust Peak-cost Constrained Reinforcement Learning Efficient policy optimization in robust constrained mdps with iteration com- plexity guarantees,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.703637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.703637Z digest=sha256:e841d948698e144b71a08de7e74bc1fb429e53b67941a148d3010110811fca98

Observation eff7d705-78de-4b06-becb-94d81c5ba068 · outbound

This paper cites Safe learning in robotics: From learning-based control to safe reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Safe learning in robotics: From learning-based control to safe reinforcement learning,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.874402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.874402Z digest=sha256:2a635b717011e5ba5e6421f552483ae078e86e64fa6407fbef97f6d1ab38c2c0

Observation 2330380a-ca15-4416-80a5-2f433f71d4fd · outbound

This paper cites Robust control barrier–value functions for safety-critical control,.

Robust Peak-cost Constrained Reinforcement Learning Robust control barrier–value functions for safety-critical control,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:42.990903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:42.990903Z digest=sha256:7eb62b334c930ed8f4890fc1288efe8e2f40e285902ff9aec6386fb6a7bdfa47

Observation 49f2e71a-0bb3-4290-b37c-6e5a6afac641 · outbound

This paper cites Learning barrier certificates: Towards safe rein- forcement learning with zero training-time violations,.

Robust Peak-cost Constrained Reinforcement Learning Learning barrier certificates: Towards safe rein- forcement learning with zero training-time violations,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.140458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.140458Z digest=sha256:a1319c7a70c0d384cbfa736de040dff4ef7304059cfd6bd49e16dc27287bd63d

Observation c71de5f5-f557-4f7f-b36c-21eb94c2f82c · outbound

This paper cites Joint synthesis of safety certificate and safe control policy using constrained reinforce- ment learning,.

Robust Peak-cost Constrained Reinforcement Learning Joint synthesis of safety certificate and safe control policy using constrained reinforce- ment learning,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.268381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.268381Z digest=sha256:9d2e24aef96dd1c8810e1e4d9fe3e0d7262f08539cb6b8fa076c9b01f0a7bb17

Observation 9af17ba4-136e-47b5-b881-3b2d2ff44d42 · outbound

This paper cites Hamilton-jacobi reachability: A brief overview and recent advances,.

Robust Peak-cost Constrained Reinforcement Learning Hamilton-jacobi reachability: A brief overview and recent advances,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.381756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.381756Z digest=sha256:928741e2aeb96e708e2a23d7f428c4f9f3908e0b59586b98f7c282ab110e0316

Observation 2a3df542-3f13-47c3-8556-376dd832015f · outbound

This paper cites A general safety framework for learning-based control in uncertain robotic systems,.

Robust Peak-cost Constrained Reinforcement Learning A general safety framework for learning-based control in uncertain robotic systems,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.452473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.452473Z digest=sha256:9962cf4d16919cb38ba07b5be1b108272137e01a180dc97a542c5fbc11d2a344

Observation 5514e28a-d393-4fd3-aab0-868757cf4e12 · outbound

This paper cites Bridging hamilton-jacobi safety analysis and reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Bridging hamilton-jacobi safety analysis and reinforcement learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.517317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.517317Z digest=sha256:cd89415ed3e632f217ed60d562bf6bc442310eaac1c2a5943d85a4770b15ad18

Observation af89d4ae-4157-470c-b02b-8fd7fe214b07 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning,.

Robust Peak-cost Constrained Reinforcement Learning Solving minimum-cost reach avoid using reinforcement learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.641773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.641773Z digest=sha256:4a07cac0cd86e9f6906659bf3346fc72872998dd24233030ceebd1bce76c8ddf

Observation 871c56dd-99b2-4a86-a578-a1921f64306f · outbound

This paper cites Con- strained reinforcement learning has zero duality gap,.

Robust Peak-cost Constrained Reinforcement Learning Con- strained reinforcement learning has zero duality gap,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.721240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.721240Z digest=sha256:102d1933b9858ca43a15cc0838e4b9f1aedf353c94ae3de346587aa7fc27cb14

Observation 99e5555c-5187-4fc8-b6ba-0544f28c96af · outbound

This paper cites Robust dynamic programming,.

Robust Peak-cost Constrained Reinforcement Learning Robust dynamic programming,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.819694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.819694Z digest=sha256:e5ea881bb2295b0946a68eea2c20243c777510cbdaac0b5c7d7b15338cf9c9b6

Observation 3ebef439-f9be-4f70-ba10-13e9297aa275 · outbound

This paper cites Natural actor-critic for robust reinforcement learning with function approximation,.

Robust Peak-cost Constrained Reinforcement Learning Natural actor-critic for robust reinforcement learning with function approximation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:43.922213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:43.922213Z digest=sha256:5cb6f5e5eb94952eb00fa1d92230322756e65461ade31480402135de2955de58

Observation 4f4e1c87-f47b-4989-bb82-cc54bf5f532d · outbound

This paper cites Integral probability metrics and their generating classes of functions,.

Robust Peak-cost Constrained Reinforcement Learning Integral probability metrics and their generating classes of functions,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:44.003201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:44.003201Z digest=sha256:621ae38d2f9c7f08f751e5e23eadc72d2b0d67184363b55b32095a16a4f8a520

Observation 25a31f43-5aaf-42f8-b33f-8d74c1ed4f0c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robust Peak-cost Constrained Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:44.081847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:44.081847Z digest=sha256:99a0d860f947698017dc9d48a87851e328e99dd37e82343beddc64c9d6b1e69c

Observation f4124912-88b4-43de-9145-a1fa65a24b16 · outbound

This paper cites OpenAI Gym.

Robust Peak-cost Constrained Reinforcement Learning OpenAI Gym

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:44.151862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:44.151862Z digest=sha256:3a3d6406392caed82c702a456268dac20899784d590ceff6a3b269d900493c44

Observation 327fcfa7-676d-4281-a973-ee01690bb500 · outbound

This paper cites Since 1−p 2 ≤ 1 2 for allp∈[0,1], this is self-consistent.

Robust Peak-cost Constrained Reinforcement Learning Since 1−p 2 ≤ 1 2 for allp∈[0,1], this is self-consistent

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T23:27:44.219249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:27:44.219249Z digest=sha256:55567e2cfcc36269ada344fe908171eb38aca9e77ad1fdeb2370e885b7cf35ee

Pith citing papers

No inbound Pith citation observations are available.