Pith. sign in

Paper Citation Record · LEDGER

Predictive Lagrangian Optimization for Constrained Reinforcement Learning

As of 12 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2501.15217.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15217 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:36:11.220159Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a851b8a-4e39-495e-89d9-e70dabceabf5 · outbound

This paper cites Human-level control through deep reinforcement learning,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Human-level control through deep reinforcement learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:11.152230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:11.152230Z digest=sha256:13b6420bb6e2ba652275e115fc0ca2e0b88a88714de18408b2b4264420c31e5b

Observation f1c9607c-fa41-443e-a16c-fa4cb1b674d3 · outbound

This paper cites Mastering the game of go with deep neural networks and tree search,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Mastering the game of go with deep neural networks and tree search,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.461677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.157730Z digest=sha256:b7806816a1fa9f459329d350999435ac762d7f662273fc3000d0d2f30096aabe

Observation 4b5092f4-f73c-4d3f-8f7e-606da9b3c761 · outbound

This paper cites Distributional soft actor-critic: Off-policy reinforcement learning for addressing value estimation errors,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Distributional soft actor-critic: Off-policy reinforcement learning for addressing value estimation errors,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.446832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.162463Z digest=sha256:27643190ecd67544bbeb8bc32215913f7d13d61ca468c16ce7e27afe6b32a771

Observation d0f7a593-2512-485b-a701-94ffbbb68fad · outbound

This paper cites A transformation-aggregation framework for state representation of au- tonomous driving systems,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning A transformation-aggregation framework for state representation of au- tonomous driving systems,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.432376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.166882Z digest=sha256:189f5039809547fc4a875af94cfa09a17507ef60fd2a718c5eef8c79535a7b28

Observation 559ca1fd-67fc-428c-98c1-0d1223bbd7d8 · outbound

This paper cites an unresolved cited work.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:36:11.417777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.171305Z digest=sha256:34baa84418934ad33b891b3045e43aa2dad9ea615df7baaa43d2b293738dafbc

Observation 2c0456ab-157a-4930-b08b-4430b6f269bf · outbound

This paper cites Direct and indirect reinforcement learning,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Direct and indirect reinforcement learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.402702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.176191Z digest=sha256:6a48984a5f9d5454e70a9416e9417e1d952f0a1fba1e1fc3440031f119a44a05

Observation eaf139ff-44c5-4d61-90e2-573c482a486f · outbound

This paper cites A dynamic penalty function approach for constraint-handling in reinforcement learning,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning A dynamic penalty function approach for constraint-handling in reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.385834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.182344Z digest=sha256:be97744fd59e3877d6c492347e06f304ef7a18fe3a93d9a328eca0d703cf546a

Observation 5e382ad8-89aa-45ab-a24c-03adecf7b045 · outbound

This paper cites Self- learned intelligence for integrated decision and control of automated vehicles at signalized intersections,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Self- learned intelligence for integrated decision and control of automated vehicles at signalized intersections,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.370392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.187789Z digest=sha256:2c8c84a8431da6e4988493a317db0232b3c794ef0199d9a74959072ff610f507

Observation 8fcd0f6a-4877-4e97-bc40-228b30594f1c · outbound

This paper cites Learning safe policies via primal-dual methods,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Learning safe policies via primal-dual methods,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.355274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.192364Z digest=sha256:4d69c7e4ee8fa25df6ebe56aa8508ff36435a51bc398b85305afe14aa5ae9de1

Observation 3adb2ecd-37e9-412e-9bc7-d9fa8bf646e2 · outbound

This paper cites Dynamical, symplectic and stochastic perspectives on gradient-based optimization,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Dynamical, symplectic and stochastic perspectives on gradient-based optimization,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.338350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.197335Z digest=sha256:159b9d7105da878a2706c28a06dd98e78d693c77ba2ab591608d32aff64735f7

Observation d2bac549-e87d-4faf-98f1-423e9b374c6e · outbound

This paper cites Responsive safety in re- inforcement learning by PID lagrangian methods,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Responsive safety in re- inforcement learning by PID lagrangian methods,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.321304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.201907Z digest=sha256:3bf6c26e27e8fef7233b165fa1e742fee99f9ab3667c882d653ec64a403a3f46

Observation 9586ec56-8c2c-4bd9-8de8-93f16ca8bca8 · outbound

This paper cites Separated proportional-integral lagrangian for chance constrained reinforcement learning,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Separated proportional-integral lagrangian for chance constrained reinforcement learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.305139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.206495Z digest=sha256:f5e7f905d248b63b145d02b914f063f1a4dfa4137632dc86f51955a2278000b3

Observation 68dee65c-b556-4932-822c-bcf042e05201 · outbound

This paper cites Model- based actor-critic with chance constraint for stochastic system,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Model- based actor-critic with chance constraint for stochastic system,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.289398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.210943Z digest=sha256:d28e152d30b5803f003e0e1f0eea1f277cf70de184d81bd38ae9c7e5537f66b9

Observation 3072ebe2-72ff-4e3a-9328-6b3c26e8080f · outbound

This paper cites Gops: A general optimal control problem solver for autonomous driving and industrial control applications,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Gops: A general optimal control problem solver for autonomous driving and industrial control applications,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.273556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.215935Z digest=sha256:e524abf8394a5931e3a36a150060aad028229603eafc69425edd445d04ae1a91

Observation d01f9777-6183-40de-89dc-611f6e010edf · outbound

This paper cites Enhance generality by model-based reinforcement learning and domain ran- domization,.

Predictive Lagrangian Optimization for Constrained Reinforcement Learning Enhance generality by model-based reinforcement learning and domain ran- domization,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:36:11.257771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T14:36:11.220159Z digest=sha256:4e52a2e192c0c702a076019a623481934f02738bc8adc54554628d0ea8df7f85

Pith citing papers

No inbound Pith citation observations are available.