Pith. sign in

Paper Citation Record · LEDGER

Operator Splitting for Convex Constrained Markov Decision Processes

As of 15 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 1 inbound Pith citation observation for arXiv:2412.14002.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14002 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:40:45.727688Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T01:40:08.230114Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T12:55:44.116744Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact4
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d5debb6-f0c9-42d9-8d71-e9b5c137e1ba · outbound

This paper cites Mastering the game of go without human knowledge,.

Operator Splitting for Convex Constrained Markov Decision Processes Mastering the game of go without human knowledge,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.474791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.523266Z digest=sha256:dcac68e0b79d368dc1320af701759dc13b0e3a790047ae91386ae4ae77fcb521

Observation a14a017c-7b4c-46bb-93e6-842395d858bc · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Magnetic control of tokamak plasmas through deep reinforcement learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.463803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.527737Z digest=sha256:05ba2021402aa4779532956cc7e0675dd97ce889ad269e226c09658727ea2ba8

Observation fc2e3662-6eba-4b80-8a36-fd86e8f01db8 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.531734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.531734Z digest=sha256:497f3ac960992632756ffd1977633efbb5879b25d6c520aff720cd4012ced91b

Observation d5ac92ef-5561-4ba3-9ad0-9402d472a519 · outbound

This paper cites Altman, Constrained Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes Altman, Constrained Markov decision processes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.535342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.535342Z digest=sha256:7795146e4a0ff8a1b23f86e8243e670ded75f7802f21eea5d5bf7ca6771f2175

Observation 5a515a8e-f66a-489c-a58b-7e289c4e5f3f · outbound

This paper cites Policy gradients with variance related risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy gradients with variance related risk criteria,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.438344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.539032Z digest=sha256:4fd896fa857ab8d97ac3cbb656497726480f7b954e47d5588a5d033190c05f48

Observation 9ec6acb3-8b02-4766-91c3-e51e28f664db · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria,.

Operator Splitting for Convex Constrained Markov Decision Processes Risk-constrained reinforcement learning with percentile risk criteria,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.426469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.542889Z digest=sha256:aab9b78807696a7a799429cdae43342001f3f5c1a62ded08c42b0a52380b6a4c

Observation ee91af8a-5923-494f-a011-6aca5f6ab9d7 · outbound

This paper cites Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,.

Operator Splitting for Convex Constrained Markov Decision Processes Control and optimization meet the smart power grid: Scheduling of power demands for optimal energy management,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.416377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.546898Z digest=sha256:f69d8e00a4233270536c8280033aeff80477f2455f04c0d24bfc03e5060b079f

Observation 542b87d3-b08d-435d-be6e-0e07797a0d49 · outbound

This paper cites Constrained policy optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained policy optimization,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.406379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.550549Z digest=sha256:a7ea3bd3115b602421e859c091edd0e289bb06352e1d9a5945521a0815073e38

Observation 45b38f5c-3b29-440f-81cb-573b57a0950d · outbound

This paper cites Dynamic programming equations for dis- counted constrained stochastic control,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming equations for dis- counted constrained stochastic control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.396802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.553944Z digest=sha256:6fe4aa2b57956d30ae0ae8a043e34c6d7da30d12a5b3337c7fb17bf230e84324

Observation ad5d1b4d-e489-46fd-b2bf-0ace75ad2016 · outbound

This paper cites Dynamic programming in constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming in constrained Markov decision processes,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.387474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.557316Z digest=sha256:8226bc85be90c8d39c59d88aeb0fea6fbfc13c5431552824e03aa538112092ab

Observation cd8afd8c-185d-45f4-bb8d-cdd29e24ae5f · outbound

This paper cites A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes A Gradient-Aware Search Algorithm for Constrained Markov Decision Processes

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.942537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.560934Z digest=sha256:c0b7bdfc537809ca58957b5edd3331982699f9b50170c446833767c1e62d6f94

Observation 81bb3222-c4a7-4975-9bfa-60d1a820a59d · outbound

This paper cites Natural policy gradient primal-dual method for constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural policy gradient primal-dual method for constrained Markov decision processes,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.376767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.564782Z digest=sha256:c09fe301a1831c772372fe123903078b13506aeb827399c50d15848adb5d425e

Observation a42db668-8b00-4eca-88ab-0631211d5210 · outbound

This paper cites Learning policies with zero or bounded constraint violation for constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Learning policies with zero or bounded constraint violation for constrained MDPs,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.366953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.568209Z digest=sha256:d25ccd1c09a1f52c65bdb87bd2b9d4959889061842e83bf1bae56f2e939d26b4

Observation e5e4fd73-681f-4d25-8067-da8a5f2ce359 · outbound

This paper cites State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards.

Operator Splitting for Convex Constrained Markov Decision Processes State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.928287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.571142Z digest=sha256:5e20548ea972bd80bc60ca559c4ae02c61c0ba5021e3d54dc942c5e81d95dc74

Observation 059a8990-aadc-44e1-b9e0-bbe49e2e94f9 · outbound

This paper cites Constrained MDPs and the reward hypothesis.

Operator Splitting for Convex Constrained Markov Decision Processes Constrained MDPs and the reward hypothesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.357018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.574448Z digest=sha256:616ec85b01e89801980abf000ab3af82fce66d88be8d6dd2eecfdb472245d3d0

Observation dc886c82-2731-4f9d-a3a9-9c3e88d9d4ec · outbound

This paper cites Two “well-known.

Operator Splitting for Convex Constrained Markov Decision Processes Two “well-known

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.346626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.576948Z digest=sha256:b1d31db31a352e4606cee3ab51da0a4cd559363dde0b77e3ffb5f6bfbae431e2

Observation 94358415-df0b-4c2c-8470-c8a9c5a99ca8 · outbound

This paper cites Algorithm for constrained Markov decision process with linear convergence,.

Operator Splitting for Convex Constrained Markov Decision Processes Algorithm for constrained Markov decision process with linear convergence,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.336444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.579723Z digest=sha256:59b570f165101b236b547c4d681510ec28d4345f2adde38c72384572f66a9802

Observation df37e178-a510-4759-9e23-c8ca8325bc49 · outbound

This paper cites Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process.

Operator Splitting for Convex Constrained Markov Decision Processes Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.912662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.582528Z digest=sha256:e0c384a917cc3a0b6ce43e57295acb2af13d1320886db4707b9affb52c5f1805

Observation 436de9d0-9ff3-4e8f-8b7a-e28e86f689d4 · outbound

This paper cites Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes.

Operator Splitting for Convex Constrained Markov Decision Processes Cancellation-Free Regret Bounds for Lagrangian Approaches in Constrained Markov Decision Processes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.585747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.585747Z digest=sha256:4d5b6365d7e341bf101221cb869ae59fbeb6bbd4eda2cf6bf7903e6a816a22ea

Observation e7dc5fbf-0b77-45e9-bc34-8fb93dbf9c2d · outbound

This paper cites Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs.

Operator Splitting for Convex Constrained Markov Decision Processes Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.589096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.589096Z digest=sha256:4fa4cea5b45ddb3728879395ef510d8aea3afd512d8ee79fc2345f4ccc09c54c

Observation d9b00ec9-7e4a-45d9-b6c9-35dac966afef · outbound

This paper cites Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reload: Reinforcement learning with optimistic ascent- descent for last-iterate convergence in constrained MDPs,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.324278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.592853Z digest=sha256:4afedb05bc3246937a558bb5671f2c85d26c9f15e6d12de4a0baab50b37d572b

Observation 15e44324-8a48-4e21-aec3-1d4ff6d76fea · outbound

This paper cites Ipo: Interior-point policy optimization under constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Ipo: Interior-point policy optimization under constraints,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.313087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.596272Z digest=sha256:fb2c04d6bd70113d534093873c0d100700830340ae451f9f74f330e6f1a9110c

Observation d9ca5f64-4792-4ad5-8b81-cca65e75d970 · outbound

This paper cites Projection-Based Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Projection-Based Constrained Policy Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.599554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.599554Z digest=sha256:5afa0b24498e39d6ffeaa2fd00bd21fa17d278406e4811cf276aca14fd3edd64

Observation 0bc2fc3e-f15d-42a0-b7ac-9ed2ce28805e · outbound

This paper cites Reward is enough for convex MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Reward is enough for convex MDPs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.302622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.604198Z digest=sha256:af637d05bcbdb9153b0dc656aebe1950df7bdd0a1761218d27a0032eec1e8ccb

Observation 3b81aac5-0a43-47d3-8586-bddbc21ce90a · outbound

This paper cites Apprenticeship learning via inverse reinforce- ment learning,.

Operator Splitting for Convex Constrained Markov Decision Processes Apprenticeship learning via inverse reinforce- ment learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.292322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.607868Z digest=sha256:b9925cb673a60b5ad40ea7f764803ddd700ad74c9758d6d2acfbd84a3e047fcb

Observation 41a27b59-5d10-424e-9af1-3f570162db0b · outbound

This paper cites Provably efficient maximum entropy exploration,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient maximum entropy exploration,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.281858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.611007Z digest=sha256:48a4f3e2c351e7da8f3d54d09f7fb22c3acd1f805a181739fbedc54e231ad6d3

Observation b9945c27-18a9-4bb5-8a98-148d484fa784 · outbound

This paper cites Diversity is All You Need: Learning Skills without a Reward Function.

Operator Splitting for Convex Constrained Markov Decision Processes Diversity is All You Need: Learning Skills without a Reward Function

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.614529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.614529Z digest=sha256:5b549d96d0e5b34ab310075d9c9f8f82468eb018361d22f1f9e491e3677c37da

Observation a405dd85-26bc-424e-9b46-9cece872a9ad · outbound

This paper cites Policy-based primal-dual methods for convex constrained Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Policy-based primal-dual methods for convex constrained Markov decision processes,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.271500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.618422Z digest=sha256:d643c5f861b2deb0436a7627847a52a1d8fce8185de9d9100a1516e33940c556

Observation 5d238fcc-519d-47b8-97fe-741bdfe0b226 · outbound

This paper cites Variational policy gradient method for reinforcement learning with general utilities,.

Operator Splitting for Convex Constrained Markov Decision Processes Variational policy gradient method for reinforcement learning with general utilities,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.260720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.621795Z digest=sha256:984e596dd66089e39c9c1ece447afa091b74075d66e047388088449e0663710e

Observation 5c4a64fb-80a3-4e21-babd-fb34144bf916 · outbound

This paper cites Reinforcement learning with convex constraints,.

Operator Splitting for Convex Constrained Markov Decision Processes Reinforcement learning with convex constraints,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.249377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.625100Z digest=sha256:13965299122f07dedc951840f5b0373d81a69760cf8fecab222f2330a6172fa9

Observation 6452146a-e617-4068-b5a7-af6bf91dbf54 · outbound

This paper cites A simple reward-free approach to constrained reinforcement learning,.

Operator Splitting for Convex Constrained Markov Decision Processes A simple reward-free approach to constrained reinforcement learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.238752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.628643Z digest=sha256:9bbd405369c5f57e1e0ea7e51db891c6ff9906d8d639742f56c68aa3b8e14241

Observation 002b4718-6274-4d5d-8a35-5130c8ea7d9c · outbound

This paper cites Bauschke and P.

Operator Splitting for Convex Constrained Markov Decision Processes Bauschke and P

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.227584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.632070Z digest=sha256:9ce763496031cff08fe958666859519f57b97144c9533d57a8af9f6e5ddac782

Observation aae25587-1f49-4ccd-bdf5-ae98d45aed63 · outbound

This paper cites Distributed optimization and statistical learning via the alternating direction method of multipliers,.

Operator Splitting for Convex Constrained Markov Decision Processes Distributed optimization and statistical learning via the alternating direction method of multipliers,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.635420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.635420Z digest=sha256:a72cb6bed6189dafdc63b1098e3d2aaf88ef1d9aacfc155807e155407f116cf3

Observation 8694d196-c0b1-433a-91ec-0f717a1e6517 · outbound

This paper cites A note on the equivalence of operator splitting methods,.

Operator Splitting for Convex Constrained Markov Decision Processes A note on the equivalence of operator splitting methods,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.210336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.638587Z digest=sha256:d121ffa229c01ccdb8a8c5bd11cabac514c658923672affd293395d2765a34b4

Observation 729dc1db-fde2-477b-9cc8-c21276720462 · outbound

This paper cites Provably efficient algorithms for multi-objective competitive RL,.

Operator Splitting for Convex Constrained Markov Decision Processes Provably efficient algorithms for multi-objective competitive RL,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.198587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.641928Z digest=sha256:39318be97d646905185bf0419b4c0d14278034c2e92d079ca134e6051f043f0b

Observation 148017e3-2317-4d51-86ef-164c340ffebb · outbound

This paper cites A splitting method for optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes A splitting method for optimal control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.187573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.645327Z digest=sha256:811da9e0b25f520d459a67d08d7e264e1becaabb3b01a257e561a00b30d7ece6

Observation 335a47a6-b86e-4328-95d5-b970d8ec748d · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

Operator Splitting for Convex Constrained Markov Decision Processes A unified view of entropy-regularized Markov decision processes

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.648460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.648460Z digest=sha256:f5836919e4af8eb8dc39b4e8a7dfd445eec9195e647a1e973c9b8bbad7996160

Observation 34500308-5375-46e9-a658-cf89b33ce04e · outbound

This paper cites A theory of regularized Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes A theory of regularized Markov decision processes,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.176808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.651895Z digest=sha256:58b173fdbededda55138142c58c13a8d463d61457386c0c533f610f095c7e11b

Observation 47314e45-c403-49b2-9885-921c406adbef · outbound

This paper cites Dynamic programming through the lens of semismooth Newton-type methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Dynamic programming through the lens of semismooth Newton-type methods,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.165676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.655087Z digest=sha256:0b9d5e5bf6292523205f397167fc28ebddac0fee61044215fcde80f695966840

Observation f226bed8-e293-433e-ae8c-f2f142ea9ed2 · outbound

This paper cites From optimization to control: quasi policy iteration,.

Operator Splitting for Convex Constrained Markov Decision Processes From optimization to control: quasi policy iteration,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.658357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.658357Z digest=sha256:6efacd27b3481f148c7a2364c1492bcbc73c52f0c89844ce0b3f5ffcea21ca49

Observation c85f91e9-7369-41ab-9e12-03177f016048 · outbound

This paper cites On the minimal displacement vector of the Douglas–Rachford operator,.

Operator Splitting for Convex Constrained Markov Decision Processes On the minimal displacement vector of the Douglas–Rachford operator,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.154673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.661823Z digest=sha256:0e134d6eec483fb042b309997f57975f94e93500d88cc06b3f940a147d5c6181

Observation 845b7a3d-417a-45ab-a633-b0261fba5ace · outbound

This paper cites On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas–Rachford algorithm for solving possibly inconsistent optimization problems,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.143746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.664885Z digest=sha256:704b600313445b2b93ded7afcc6a148ff1d633d7a38d9d843abc10d6a794d43c

Observation d1af1847-d264-4ecb-b59f-461b404f105a · outbound

This paper cites Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in alternating direction method of multipliers for convex quadratic programs,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.133546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.668212Z digest=sha256:f557a132770b9ed302492a152c51c6ee320ad62c23c324f24c9abea26b94140e

Observation 38476989-5207-4d52-9083-c9d5b7edbc04 · outbound

This paper cites A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs.

Operator Splitting for Convex Constrained Markov Decision Processes A New Use of Douglas-Rachford Splitting and ADMM for Identifying Infeasible, Unbounded, and Pathological Conic Programs

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:40:45.774770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.672091Z digest=sha256:a3baff52611541c6d9961ef0781e44e83a3d7e2ecaaa166b6bf1f780be0b97e0

Observation c19741f1-0603-4eff-8ff8-a0f75088959b · outbound

This paper cites Infeasibility detection in the alternating direction method of multipliers for convex optimization,.

Operator Splitting for Convex Constrained Markov Decision Processes Infeasibility detection in the alternating direction method of multipliers for convex optimization,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.122245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.675613Z digest=sha256:7f082598e4ce9720dd7061dd925afee38b75bd329ef3c785f1089854ef3382b2

Observation 2d88bf23-f813-4fe2-943b-38f62e751c1b · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T12:40:46.108724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.678725Z digest=sha256:e1b1b4b21612c0519c41db8ffc4dd251b0b467be99a7d1f575d5c4b9a113a1a4

Observation 1b6545ba-014c-4585-a7dd-adf545063e20 · outbound

This paper cites On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,.

Operator Splitting for Convex Constrained Markov Decision Processes On the Douglas—Rachford splitting method and the proximal point algorithm for maximal monotone operators,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.096953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.681633Z digest=sha256:f704f91c72ee8db24283237ec92679c54b675fcf80201b168a3dbcefe505ad02

Observation c69d66ec-44ed-4ae4-aa7c-882db7784a41 · outbound

This paper cites On the convergence of the coordinate descent method for convex differentiable minimization,.

Operator Splitting for Convex Constrained Markov Decision Processes On the convergence of the coordinate descent method for convex differentiable minimization,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.084852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.684600Z digest=sha256:13aab8d5098678822b2eb7b0f1ab40300a9919459a5efd0347523adb873e422f

Observation b49bbd16-32d3-4d0c-a4da-e95a873a9931 · outbound

This paper cites Nocedal and S.

Operator Splitting for Convex Constrained Markov Decision Processes Nocedal and S

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.072601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.687334Z digest=sha256:cdcb1b8719e30da2d17f9c514bdd56df7481cd1d82df1a0597e59bf1492b4843

Observation 9e4d7071-9aac-4c60-99b2-5ec01d98fb67 · outbound

This paper cites Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,.

Operator Splitting for Convex Constrained Markov Decision Processes Operator-splitting methods for monotone affine variational inequalities, with a parallel application to optimal control,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.060999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.690317Z digest=sha256:0b4eb817a76813baa1d373da8b423fc05a2c5d3bcf5e87fe06dab1b0d3308ce1

Observation a1b6c6ae-34c2-4999-90b9-0e8a26dfd4b6 · outbound

This paper cites Parallel alternating direction multiplier decomposition of convex programs,.

Operator Splitting for Convex Constrained Markov Decision Processes Parallel alternating direction multiplier decomposition of convex programs,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.693417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.693417Z digest=sha256:469c392880128248a4a7ea41f4051873938f18ba4f2d32984ac1c72d63ad37ae

Observation 5dbb4e01-7ed9-4674-8b68-5ef2c559d5f2 · outbound

This paper cites Natural Actor- Critic Algorithms,.

Operator Splitting for Convex Constrained Markov Decision Processes Natural Actor- Critic Algorithms,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.041883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.696698Z digest=sha256:2c87d53c9acccce9887899401761b53d2ef185f75012159b684752f3bda91d84

Observation c93d3625-2ce4-4e7d-a5b2-e990d6dbc021 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Operator Splitting for Convex Constrained Markov Decision Processes Pytorch: An imperative style, high-performance deep learning library,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.699652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.699652Z digest=sha256:0e06258a94241bd00f3f7a2fc7e58a6a1f940e64d0dc3e6a5f32a4c918122425

Observation 88f42b45-ae06-4a60-8f92-525e4980a090 · outbound

This paper cites PID accelerated value iteration algorithm,.

Operator Splitting for Convex Constrained Markov Decision Processes PID accelerated value iteration algorithm,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.022465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.702427Z digest=sha256:2005f2da25822f5310e25cfdb4f153c4a6819905ffc7530e21e3e0dd01f6d8f8

Observation e5599083-65e5-438d-9d0e-c179ab9b09e5 · outbound

This paper cites Scalable first-order methods for robust MDPs,.

Operator Splitting for Convex Constrained Markov Decision Processes Scalable first-order methods for robust MDPs,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:46.008711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.705410Z digest=sha256:2d35a9ee75d43d6026d1202456430fd091a129ae947a8c20d871bb62bff763eb

Observation 39b06f7f-f487-4906-9d07-fdd4f6b374f1 · outbound

This paper cites Integrating a partial model into model free reinforcement learning.,.

Operator Splitting for Convex Constrained Markov Decision Processes Integrating a partial model into model free reinforcement learning.,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.996527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.708307Z digest=sha256:517046010b95514209dbd71c23b89fb2f5a59414a7943d19cd978abd6634e023

Observation 63fdc30f-6f3c-4af2-a99d-59108ef8521e · outbound

This paper cites Gurobi Optimizer Reference Manual,.

Operator Splitting for Convex Constrained Markov Decision Processes Gurobi Optimizer Reference Manual,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.711404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.711404Z digest=sha256:a69dcec3c89bb9a96a9118e3463b6b22a010ca1b5c5207949a2a2d68b64543e8

Observation 259e0a59-fbdb-4c88-a3c2-6483e6ec897c · outbound

This paper cites Safe policies for reinforcement learning via primal-dual methods,.

Operator Splitting for Convex Constrained Markov Decision Processes Safe policies for reinforcement learning via primal-dual methods,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.978868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.714654Z digest=sha256:9d8accb549ce5f005e39a8f681f216f62971a04adb2b4732385b2bb0f2df2417

Observation 2729d892-f1e5-4e18-b5ff-3355062fa1bd · outbound

This paper cites Conic optimization via operator splitting and homogeneous self-dual embedding,.

Operator Splitting for Convex Constrained Markov Decision Processes Conic optimization via operator splitting and homogeneous self-dual embedding,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T12:40:45.968594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T12:40:45.718057Z digest=sha256:35d3396b31a9fb3cae33006c8000a65a5183d5f64a83967a8c526b6f687e3245

Observation d61bd2c1-8f1d-44a1-b1e3-8e9084a80edd · outbound

This paper cites Reward Constrained Policy Optimization.

Operator Splitting for Convex Constrained Markov Decision Processes Reward Constrained Policy Optimization

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.721211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.721211Z digest=sha256:e385b2905f65b3f0718e6884683ea18dd99c1b6cb75330c141fe1f2c9d3959c1

Observation 992f0858-c3e3-459b-ab05-cc8006d66e68 · outbound

This paper cites Markov decision processes,.

Operator Splitting for Convex Constrained Markov Decision Processes Markov decision processes,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.724621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.724621Z digest=sha256:2fc641b1bd6578972d7738e4ba6379b8911957dcf1ae6977a113c400c5030c42

Observation 32b6a08f-3524-4360-8b11-de9e3173d032 · outbound

This paper cites an unresolved cited work.

Operator Splitting for Convex Constrained Markov Decision Processes Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.727688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.727688Z digest=sha256:9ffd193b4e380fc9597210a3e8ce582814c177028c34dc37d98bdf48d883ada1

Pith citing papers

Observation c91b5c64-12f6-42f0-b676-61bc626b43da · inbound

Joint Chance Constrained Safe-Optimal Control cites this paper.

Joint Chance Constrained Safe-Optimal Control Operator Splitting for Convex Constrained Markov Decision Processes

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T12:55:44.118969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-01T01:40:08.230114Z digest=sha256:a26751b1d80e9d9443c2984329b4019134b733fc710bc6cd280adec934b2522a