Pith. sign in

Paper Citation Record · LEDGER

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form

As of 5 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 2 inbound Pith citation observations for arXiv:2408.16286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.16286 v5

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T21:58:56.180393Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T10:50:40.841967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T08:49:42.710069Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact22
  • verified fuzzy63
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 129e1ef9-7bc7-492c-a1a0-d2e87f4ea2d8 · outbound

This paper cites write newline.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.979602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7a147793c99d58943242a0ef2b89e85caa50e7fa984cb9704523bc1385da801f

Observation dc78442c-4ca0-4f03-bd47-5516f17fe567 · outbound

This paper cites On Frank-Wolfe and Equilibrium Computation.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On Frank-Wolfe and Equilibrium Computation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.012306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6ccb748293850e36566978a8d3da3d46a4aa9abe0d5d27c6f16342597c7a3820

Observation 47932635-675f-48b7-9383-a37d031108ec · outbound

This paper cites Constrained Policy Optimization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Constrained Policy Optimization

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.962432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:ebedea49301f82cf22b775e167b92c6dca8e5099ce96d766c91659b5dee7dbf1

Observation 2a52e389-5b95-4acc-a539-ebe9c64e9baa · outbound

This paper cites On the Theory of Policy Gradient Methods: Optimality, Approximation, and Distribution Shift.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On the Theory of Policy Gradient Methods: Optimality, Approximation, and Distribution Shift

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.072439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:0868c3b7073540fbb1f7f0ba9ceb83b435827a6d7356ebb122b28de0bf777ffa

Observation a9a7c2ae-ad56-4b44-8d54-bc41cc35ffcf · outbound

This paper cites Constrained Markov Decision Processes , volume 7.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Constrained Markov Decision Processes , volume 7

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.992363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:beb92e6e54799e9eadc5dbb9d6bc62f6d094e4afe953aedcaa0eb9b4e7a072ee

Observation 93f3cc20-6c13-4a5b-aba1-6b739e2f4939 · outbound

This paper cites System Level Synthesis.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form System Level Synthesis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.982573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:e945fdccb33f1321c52dd899c966abe5c3f87f63a440ee402dd28020f2da4001

Observation 229ec244-9919-4037-b280-6c692b566800 · outbound

This paper cites On the Generation of Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On the Generation of Markov Decision Processes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.941185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7d7d0c13ac9cafe0c0cde9365593c893638652c09fae245fcb70410214153fac

Observation 3f5d0d96-6153-4d54-93cc-88328e41525d · outbound

This paper cites First-Order Methods in Optimization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form First-Order Methods in Optimization

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.944576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:aaece483ec1a11ab1a3a9d9fd164ce0601dc12808f6569e5be52f77940c350cd

Observation 3803a6e2-35c6-4cc6-9f6c-d5e42ddd4378 · outbound

This paper cites Bellman, R.E.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Bellman, R.E

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.948578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:03131608866c5dcab8003902b7a5a46a99ed8f7c1aedb4f616fd65c87b5b6bf5

Observation f051fdb7-c527-4d86-9ef4-17a696393a2e · outbound

This paper cites Robust Model Predictive Control: A Survey.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Model Predictive Control: A Survey

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.079204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:ae30b12e869d0a37a56c74bca168a5ee43b28bc42074f514b8208b3d1f7f8f6c

Observation 3aae35ee-d4f5-45bb-9da1-f4ef6f852307 · outbound

This paper cites Robust Optimization--a Comprehensive Survey.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Optimization--a Comprehensive Survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.075675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:9ef6a8134fb4bdf15668e0b9d52ce9ec4a698ca1e13e04d508e068b021217e84

Observation e0f07133-47ad-4ea9-8f15-d9fc8578e453 · outbound

This paper cites Convex Optimization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Convex Optimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.968462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:fd612e71908b9d404ba2802280a25b472676facaef21e60358b40249370f73f3

Observation f13d1aab-de65-4901-a54f-335633b034d4 · outbound

This paper cites DOPE: Doubly Optimistic and Pessimistic Exploration for Safe Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form DOPE: Doubly Optimistic and Pessimistic Exploration for Safe Reinforcement Learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.032466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:8d0a1186134b7d8716c9802e64184815ec68cd86be3172b1a25e71cf5d89a32e

Observation 34e346e0-cd9e-4201-b4fa-db63b0665bcd · outbound

This paper cites Approximate Constrained Discounted Dynamic Programming with Uniform Feasibility and Optimality.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Approximate Constrained Discounted Dynamic Programming with Uniform Feasibility and Optimality

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.987122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:3e489cf21d9d1d04da39448985c34dd3ad0a73931433cfedf596907cdea7e834

Observation 768ee083-fc5a-4a84-b234-94f1954003d8 · outbound

This paper cites Dynamic Programming Equations for Discounted Constrained Stochastic Control.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Dynamic Programming Equations for Discounted Constrained Stochastic Control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.976367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6e950a3ddfda43dea1574d46dc9214b697454ecbfa3ba1a35222524d3eca0d5a

Observation e5721103-dc34-4027-a18f-0fa2d1373cb3 · outbound

This paper cites Non-Randomized Control of Constrained Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Non-Randomized Control of Constrained Markov Decision Processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.965526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c3eb695834a19a5fe293eb09687d3a1eb8c7150c49577434d6e69f7b7d04ab68

Observation b13a324d-039d-4ade-999f-cc86ce9fe26b · outbound

This paper cites A Primal-Dual Approach to Constrained Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.982953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7893870cd06f13c058e70b349ddd284c936ee78f712939ec27e4ed483692590d

Observation 0a1ca860-4511-4316-b725-d1f7788ce0dd · outbound

This paper cites Distributionally Robust Optimization for Sequential Decision-Making.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Distributionally Robust Optimization for Sequential Decision-Making

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.035949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:0daca3bdaa4bbc803457d6dd1264af62f62c63302950895120e363e1d2dd8cd7

Observation 6d3659c8-bf79-4b52-96ef-99a10a134986 · outbound

This paper cites Nonsmooth analysis and Control Theory , volume 178.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Nonsmooth analysis and Control Theory , volume 178

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.999288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:fc5959733ad392eb3baee09a0a6b0d4c1ef7644d7f48f772e238c1ae2db5a647

Observation c93edf99-c46e-4f92-be99-d5a2ffa67fdb · outbound

This paper cites Optimization and Nonsmooth Analysis.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Optimization and Nonsmooth Analysis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.032583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:73f99c6368b528db4eb26271e9ffdd26b9f7269bcce78df4d07c5bb6470d68f4

Observation cfd284a9-4453-4e48-8933-733180bc8334 · outbound

This paper cites Towards Minimax Optimality of Model-based Robust Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Towards Minimax Optimality of Model-based Robust Reinforcement Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.949925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:2291fb4dec2a81bb5b6de348844902a4d38b54e22b31f9e71ba945ed272d329f

Observation 1b65327e-99f5-4da8-8895-4e1f7fc2f80f · outbound

This paper cites Unifying PAC and Regret: Uniform PAC Bounds for Episodic Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Unifying PAC and Regret: Uniform PAC Bounds for Episodic Reinforcement Learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.028862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:9e5aced9c1ef60149b26f8e341859fc767be49be99ea0464ac5ebacfe345d73d

Observation b8d59e85-83ef-434c-8063-dc18c4674480 · outbound

This paper cites On Linear Programming in a Markov Decision Problem.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On Linear Programming in a Markov Decision Problem

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.024720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:1ec306bc0d533e2a7bc051e4874a13fc696eee9968c88454245267fad55660b2

Observation 92d598cb-8dc7-4f4f-b045-b6e1044f506b · outbound

This paper cites Twice Regularized MDPs and the Equivalence Between Robustness and Regularization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Twice Regularized MDPs and the Equivalence Between Robustness and Regularization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.055446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:084e992affa09e84ec65ba9b3d70b76a4865b23947a4160c02db0b769bc81808

Observation 02607cfd-3fd6-4583-a005-eb0370df6532 · outbound

This paper cites Natural Policy Gradient Primal-Dual Method for Constrained Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Natural Policy Gradient Primal-Dual Method for Constrained Markov Decision Processes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.021714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:65dff6fcc254242fb42140be76d73c2f60e4ab0d7bca75698783be540e8cf2d0

Observation 8ea81958-9a30-4976-89ce-daac2c21e3de · outbound

This paper cites Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.018667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:90bab2a99a6700c7fefb517191048336335f970668a81cbabd845ccdc7673384

Observation 82e67b61-a73e-4cc4-95eb-360151defd35 · outbound

This paper cites Analysis of Feedback Systems with Structured Uncertainties.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Analysis of Feedback Systems with Structured Uncertainties

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.015731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:98d91a22046dd8cafe7fc115287155c628491b5ad515a8dfcbe7562e954a07d1

Observation 4798f791-03c7-4f9b-89c6-470c8a46fc33 · outbound

This paper cites Bilinear Classes: A Structural Framework for Provable Generalization in RL.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Bilinear Classes: A Structural Framework for Provable Generalization in RL

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.089338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7a6f29b0a26a54cf85d1ed912b685d6cf83315eb861ba5f9a1894fa60a6ce31e

Observation 4f73b7d1-22f9-44c1-ac77-932e73cd1db8 · outbound

This paper cites Robust Nonparametric Regression under Huber's $\epsilon$-contamination Model.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Nonparametric Regression under Huber's $\epsilon$-contamination Model

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-23T22:03:30.945007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:2025f795dcb65a2f7665e844d43e9c402b97eb4c081133443b6500cb24d9f535

Observation e715c81e-a517-4504-8d2b-dd23015a0d4e · outbound

This paper cites Exploration-Exploitation in Constrained MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Exploration-Exploitation in Constrained MDPs

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.941084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:92543909e6d5cde9fd11f2549534832b7b53374709a764d4e350fcc6008bf3a3

Observation 71699e00-a72a-4fd2-be3c-40fc4f8027ce · outbound

This paper cites Sample Complexity for Obtaining Sub-optimality and Violation Bound for Distributionally Robust Constrained MDP.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Sample Complexity for Obtaining Sub-optimality and Violation Bound for Distributionally Robust Constrained MDP

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.012328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:06748f5064d85ed54ad0fe988808311bb6d47000b424f7771478ffdd716855af

Observation f5733f09-06dc-40d9-a3a6-f5743b8056de · outbound

This paper cites Robust Markov Decision Processes: Beyond Rectangularity.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Markov Decision Processes: Beyond Rectangularity

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.008639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:668fd7f8255bfc9e4718b21e11439c2afd585fbcc16a9f266c76e178327f103c

Observation 190d35a3-3571-4a79-9341-6250fbbdb226 · outbound

This paper cites Scalable First-Order Methods for Robust Mdps.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Scalable First-Order Methods for Robust Mdps

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.004614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:a6e062124ce07c142c22d85793cfd68710928272d2431dd21be11100edefe7e2

Observation bdd12b3c-58b4-4833-8d84-be5abab7005a · outbound

This paper cites On the Convex Formulations of Robust Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On the Convex Formulations of Robust Markov Decision Processes

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.059137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:35d6379615b5f193480039dae25154abd7ac083f12c816370a31b53eb9e463ae

Observation 38bc00aa-b505-401d-b335-39736adcb8f7 · outbound

This paper cites Learning with Safety Constraints: Sample Complexity of Reinforcement Learning for Constrained MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Learning with Safety Constraints: Sample Complexity of Reinforcement Learning for Constrained MDPs

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.000962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:1c8760ea75b11882773be093aead18f70ae7b6eead4c5125fc60cfdf8389336c

Observation cf52c55b-69c3-4d4b-82b0-545ef32c44e3 · outbound

This paper cites On Constrained Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On Constrained Markov Decision Processes

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.086031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:9d74d41b3f34b2bcc57c684072064b25c071c1c51be0f07cfd30d83cfd0f04a8

Observation 49ffaac0-d10f-47b6-a31c-892c8b44808b · outbound

This paper cites Partial Policy Iteration for L1 -Robust Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Partial Policy Iteration for L1 -Robust Markov Decision Processes

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:49.997464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:9bc60725268b7560185af2d74fb0101ce3581098b8212cd23585bf8567d6da39

Observation c90e0f95-f5e1-4963-8296-38ff05ae89a4 · outbound

This paper cites Robust Dynamic Programming.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Dynamic Programming

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.934679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:258c5360b129e56a79b396aef75b115637e7c7fc0ef5e85b39edb043dc17df8d

Observation 93218b0d-77b8-4621-8ed2-f8b5b9573764 · outbound

This paper cites PAC Reinforcement Learning with an Imperfect Model.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form PAC Reinforcement Learning with an Imperfect Model

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:49.994044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:41d43929c258e4c607875107e84f2091d4414a9f71a1b2048075f51efabe781c

Observation 9b968be1-f7cf-4e00-9447-0b12d7bb831d · outbound

This paper cites A Policy Gradient Primal-Dual Algorithm for Constrained MDPs with Uniform PAC Guarantees.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Policy Gradient Primal-Dual Algorithm for Constrained MDPs with Uniform PAC Guarantees

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.959357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6be8c8237fd632e637eb8891713ee6afc4d36de7ff7616eda2a82b7ff5266117

Observation 7fb6ef4a-e93d-48ae-9103-91bf64ed2c44 · outbound

This paper cites On Fr\'echet Subdifferentials.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On Fr\'echet Subdifferentials

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:49.990504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:857ea5925ec5cea7fa1987ae7dfde9816d7a86fd69877285deef54527ba7dc12

Observation d1b634bc-3c8c-4742-8916-ec58510bf5a2 · outbound

This paper cites Efficient Policy Iteration for Robust Markov Decision Processes via Regularization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Efficient Policy Iteration for Robust Markov Decision Processes via Regularization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.936279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:78d7a20ffa2a84be3f7b218ae69883d889ca131ff89435b4d0ffddb5d5a30169

Observation e2894738-deb9-4012-851d-09c5320ae655 · outbound

This paper cites Policy Gradient for Rectangular Robust Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Policy Gradient for Rectangular Robust Markov Decision Processes

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:49.987170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:f1811ef4546f7d23b8e7023c61653ab52b929a9eb9db428a2273d17a520816d1

Observation af79b652-80e4-4edc-ade9-3ea50c8045a2 · outbound

This paper cites Batch Policy Learning Under Constraints.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Batch Policy Learning Under Constraints

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.937936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:96ad5ee26e4cefa242b6d282d1fc934133792786311215957c749cbbc6b18847

Observation 66df72e6-34ab-4dd9-b302-10cda443b53a · outbound

This paper cites Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:49.983924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:537167d74fbfd6915468b1f5128354618f748c8547d2f397369997851bbb096b

Observation ad6ae68f-dc42-47b6-84b5-6ddabd41e6d9 · outbound

This paper cites First-order Policy Optimization for Robust Markov Decision Process.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form First-order Policy Optimization for Robust Markov Decision Process

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.920726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:56977dbda695253040549caf09ec67907d5d321f041f30e4f57b19f9d835a460

Observation fd96314a-8a8a-4d8b-8fe4-0f5d2c681a29 · outbound

This paper cites A Single-Loop Robust Policy Gradient Method for Robust Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Single-Loop Robust Policy Gradient Method for Robust Markov Decision Processes

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.968838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:ae32796f55dd5b0c654a739fd2f9fbea07b8332afd6102b136e64a411b72f9ad

Observation dee680ee-99c5-4596-af3a-827535095b74 · outbound

This paper cites Learning Policies with Zero or Bounded Constraint Violation for Constrained MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Learning Policies with Zero or Bounded Constraint Violation for Constrained MDPs

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.989125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:dcb66f7f48a9a545d9915a63349e97e54c70c82549c6cbbf8d8208e92d78676d

Observation bedf5b64-3845-4897-b77b-b5de3e5e53b0 · outbound

This paper cites Robust Entropy-regularized Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Entropy-regularized Markov Decision Processes

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.916531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:8236a658919cf7ab37cf0b37bd40478c2acf8b3e9385dc9ab33bcab1636c482f

Observation 922c77d1-352e-44fb-bc49-6baa636530d8 · outbound

This paper cites Robust Constrained Reinforcement Learning for Continuous Control with Model Misspecification.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Constrained Reinforcement Learning for Continuous Control with Model Misspecification

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.912566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:3f0ae63391df4c6eb86addd0fefb1e450504749529dcd1dd575d945e28ba5b2a

Observation 50af8e2a-640e-4ec2-aaf3-1f5110ff2db5 · outbound

This paper cites On the Global Convergence Rates of Softmax Policy Gradient Methods.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On the Global Convergence Rates of Softmax Policy Gradient Methods

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.985929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:77fcdbefda7fcdecd5f0e7559f845fdd4fd47fc2fd62b1349fea51b7831986e8

Observation 2f55bbe6-c3c2-434c-aefc-cd3309bfce24 · outbound

This paper cites Methods of Nonconvex Optimization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Methods of Nonconvex Optimization

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.907813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:203c36010b14f11b21af9ac8943b7d527025f354066cb13040b6da980f2455b8

Observation bb24d145-ba48-44f7-bd75-c440d9d18c61 · outbound

This paper cites Reinforcement Learning with Convex Constraints.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Reinforcement Learning with Convex Constraints

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.035862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c47710e4ed86a4cdffe5c3da1d8ca879a4003111d7f8c5ba3a450f0cd82e4a3c

Observation d34578d6-75b8-4280-80bb-709383bde760 · outbound

This paper cites Human-level Control through Deep Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Human-level Control through Deep Reinforcement Learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.045465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:1395f1122141ec0e74bd25aada7b08e1e42c08cb3daec6b41c5aa5a0d44a9069

Observation 0b112a79-ae19-4585-ad14-1282813a7f50 · outbound

This paper cites Truly No-Regret Learning in Constrained MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Truly No-Regret Learning in Constrained MDPs

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.898987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6460dba4160e1e958e506883bf8b8f245133fce6fa275f2bf912a847ba9225fd

Observation 94cc27a9-5b6d-4205-bbb2-c1b2d4ff155a · outbound

This paper cites Reinforcement Learning via Fenchel-Rockafellar Duality.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Reinforcement Learning via Fenchel-Rockafellar Duality

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.932264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:f48ee5d6ae8aa1036f05a976228ea302e178cb144ee6b9c652ae67ee249fa2c3

Observation d2816846-5795-4640-96dd-c7b380fa72e5 · outbound

This paper cites Robust Control of Markov Decision Processes with Uncertain Transition Matrices.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Control of Markov Decision Processes with Uncertain Transition Matrices

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.951981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7d3bb9c465e654e36a1de4d1b95c23162b6f915fd9389a4e2f735cff8e08f971

Observation af98bccf-71c2-402c-92e6-1549d6312eee · outbound

This paper cites Sample Complexity of Robust Reinforcement Learning with a Generative Model.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Sample Complexity of Robust Reinforcement Learning with a Generative Model

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.028763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:a896fb939f2bbe716ab854095691c02b82f645d068a87f19b597f467ef8995e7

Observation f45647cf-0c36-4662-8ad3-10c41ab25a3b · outbound

This paper cites Proximal algorithms.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Proximal algorithms

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.971733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:8747c9011bb73c4a3ff712a78442fc77ea1fb8624d2f178fe9122b76da4b54a3

Observation 9d928f18-f646-42eb-81da-16a3c42c1a05 · outbound

This paper cites Constrained Reinforcement Learning Has Zero Duality Gap.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Constrained Reinforcement Learning Has Zero Duality Gap

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.039014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c821177341448766ad35d6cb13cf91b2a43354c595a61389443bf4daf12abdd0

Observation 5ab447bc-8c44-4914-97ac-99c5b42e0901 · outbound

This paper cites Safe Policies for Reinforcement Learning via Primal-Dual Methods.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Safe Policies for Reinforcement Learning via Primal-Dual Methods

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.959407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:41461581187d33ba720aad20c9cee3948ff1d52ee4785df9b990f980329c22cd

Observation 8d7d9546-1389-4c68-9eed-cdd260e13aac · outbound

This paper cites Safe Policy Iteration.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Safe Policy Iteration

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.005968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:29200d80155cc794f5cd6684798642d77b0e7d73c493b1d725663ecc83cdf668

Observation 8370ecf0-345b-45d3-a46c-1554ab0623e6 · outbound

This paper cites Distributionally Robust Optimization: A Review.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Distributionally Robust Optimization: A Review

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.926217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:d4e164f2647415973b10214c8c0b3e1e680deda3eac47549c73535e6005d42eb

Observation e3b06707-cdeb-4798-bf15-0e2a2bf8b73b · outbound

This paper cites Variational Analysis , volume 317.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Variational Analysis , volume 317

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.022088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:2c96c15ab2a8a1bb98215237f5ec6a91862bf7a57e920c128051e84666120eeb

Observation b6e106d1-da69-469b-ba45-934f39ea08c0 · outbound

This paper cites Robust Constrained-MDPs: Soft-Constrained Robust Policy Optimization under Model Uncertainty.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Constrained-MDPs: Soft-Constrained Robust Policy Optimization under Model Uncertainty

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.903662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:e28a1dfe817474b43240d7090e03231f81899c795d75da1c6f78a5b60eda2e30

Observation 4709763d-2593-4b28-ba0f-e238b6df8283 · outbound

This paper cites On General Minimax Theorems.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form On General Minimax Theorems

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.062816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6ec1031f4bd7b9fd6cedbbce8f0f975021731f89441f5ab2d718f3ac44cf1da3

Observation 83b5a19d-d3db-4c0b-9ba3-fd38dcd2d209 · outbound

This paper cites Solving Stabilize-Avoid Optimal Control via Epigraph Form and Deep Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Solving Stabilize-Avoid Optimal Control via Epigraph Form and Deep Reinforcement Learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.995642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:b4728e2b196a32ecaba7489fa525bb117cd11491cff7e490e87f01bd38dcad0f

Observation 61bbc45d-2ea6-4e45-a06c-665c0a9964a0 · outbound

This paper cites Solving Minimum-Cost Reach Avoid using Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Solving Minimum-Cost Reach Avoid using Reinforcement Learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.082673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:8caaf9f6e5a7eb817974c3404ed2e65cbd7c90a4b39249598f45f9d19e2c243f

Observation aa7ad849-f2a4-40a3-a822-d6e303af089e · outbound

This paper cites Constrained Reinforcement Learning Under Model Mismatch.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Constrained Reinforcement Learning Under Model Mismatch

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.978455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:e74328fe2763523dce4ac194b796985077a473564a6fb9fac1bb420b3afd2d82

Observation 7053fabc-25a0-4eb7-a42e-1210694cc7f5 · outbound

This paper cites Reward Constrained Policy Optimization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Reward Constrained Policy Optimization

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.065868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c48717830af6491b73801d1fa37d76322788d6560cbda03dd03f87c2c2734ac0

Observation bf59b2d0-d1ec-4164-a0e6-c3c253cc45f0 · outbound

This paper cites Policy Gradient in Robust MDPs with Global Convergence Guarantee.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Policy Gradient in Robust MDPs with Global Convergence Guarantee

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.042401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:7b8afacac298b7fe2177a74e2a916a2fdca1dd46d11a2457e1257b8e60335402

Observation 4ca009c7-8fda-40f3-bbf4-40845646dcdc · outbound

This paper cites Online Robust Reinforcement Learning with Model Uncertainty.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Online Robust Reinforcement Learning with Model Uncertainty

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.009160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:d668af6df5f7f2fb73c12b23085b29e14df3eca2fc7ea2434f3c8ff0ecfc0146

Observation f357eaf3-36c2-4c49-b676-82aae3be6ced · outbound

This paper cites Policy Gradient Method for Robust Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Policy Gradient Method for Robust Reinforcement Learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.069063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:759eb21146e18e7310124e67d7fc2044b4c4c5af3cb44c19797c1c556a53113a

Observation 1df4ecea-cd46-421f-98dd-7165b08b020f · outbound

This paper cites Robust Constrained Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Constrained Reinforcement Learning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.895012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:9cfb63a1ff69cfc703a88f4c7b8b35626ac61c966998196a653e8cd86c89bf97

Observation 5378cdf3-dd43-45ed-9659-d97273c92922 · outbound

This paper cites A Provably-Efficient Model-Free Algorithm for Constrained Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Provably-Efficient Model-Free Algorithm for Constrained Markov Decision Processes

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.954711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:68882c11d1471cf7a9b0a820fa305cc65f37e6db967621fc65cc7eda3c478757

Observation 4738e297-0b61-4de9-bbbd-3c25b96c8fb6 · outbound

This paper cites Robust Markov Decision Processes.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Markov Decision Processes

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.048691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c7b31d9a74f53da9766c4d60b571c55510552802150b25382b66042dbbfb82bd

Observation f961d062-3d3a-4673-9cd3-7cb53cf5d8a7 · outbound

This paper cites Toward Theoretical Understandings of Robust Markov Decision Processes: Sample Complexity and Asymptotics.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Toward Theoretical Understandings of Robust Markov Decision Processes: Sample Complexity and Asymptotics

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.052046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:81fe9a9d9a3e825503906c9baaefbb78f67386e1464ab27775bd9e0a8b49ead1

Observation 904566eb-d44f-49e5-89f9-d9408f108001 · outbound

This paper cites Robust Markov Decision Processes without Model Estimation.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Robust Markov Decision Processes without Model Estimation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.890063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:f0c5cb685208a5e3e3f48ebb6ba452b791c28ba2549f02e12b9784bcdbbf588d

Observation 7922776d-1575-498d-9fa8-5b54ea2f4e6c · outbound

This paper cites A Dual Approach to Constrained Markov Decision Processes with Entropy Regularization.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Dual Approach to Constrained Markov Decision Processes with Entropy Regularization

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.025291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:6081511aca29dd04509554be4e8359bcdb7de9ebe818a389b7af14a205315f65

Observation 13626867-78af-47f2-b0ad-cdda4f8fc04b · outbound

This paper cites Reward is Enough for Convex MDPs.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Reward is Enough for Convex MDPs

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.018639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:f1c42f13353f835d1b58e0f5086e1f7213275dc562859492cec389eae5d17723

Observation 1d6bfaef-4e0f-463c-be4f-7355017e1334 · outbound

This paper cites Feedback and Optimal Sensitivity: Model Reference Transformations, Multiplicative Seminorms, and Approximate Inverses.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Feedback and Optimal Sensitivity: Model Reference Transformations, Multiplicative Seminorms, and Approximate Inverses

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:58.955755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c83dabdde296f860652d24c60bd06945bb507cd5a82f556ef7970a3a742d2ff9

Observation 0b6c0030-1532-494e-b032-dfeb7dfbad91 · outbound

This paper cites Distributionally Robust Constrained Reinforcement Learning under Strong Duality.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Distributionally Robust Constrained Reinforcement Learning under Strong Duality

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.964295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:e163471c38a873909f23e335ed5c6dd5dd05d5e984e9d66df31f084b070ef967

Observation df2500fa-80ac-4042-b7e0-986959c16af2 · outbound

This paper cites Constrained Upper Confidence Reinforcement Learning.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Constrained Upper Confidence Reinforcement Learning

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:15:59.003266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:ae067a5d4698da0287a7f37264def04a204eef4198734c6fef8d5b2c1fb89494

Observation ef2c8820-73ef-45e4-8ccc-e5440ae53d5a · outbound

This paper cites @esa (Ref.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form @esa (Ref

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T03:08:50.039368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:3f98420e739ee316db113bcac47960e5fe8b9d7dcc92371a07da7dcb69031a75

Observation afad99e6-166d-40c3-aa7b-9079e5873e01 · outbound

This paper cites an unresolved cited work.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-05-24T03:08:50.043144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:60e2de30cc77f35af771ddecce53e83b152188e82a516c4fd1684e4ac1766ca6

Observation cb767936-0230-4834-a91e-20b79afa7ac7 · outbound

This paper cites Victoria Beckham.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form Victoria Beckham

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.973878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:2ffb26a7ed7c216ac7d5299624e027a499fe5e9bf348e2d7f14d9189c078cd1c

Pith citing papers

Observation aa1eb4c0-12a7-41dc-9a67-b4d2cdd38fbf · inbound

Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees cites this paper.

Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:40:27.187581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:36:59.355147Z digest=sha256:f8bb19fc7246be06d429334f384e27cbbdeace2feadb5161830a32bb21ef09f1

Observation 390cb399-72b5-40ff-8cdd-1ee17966621c · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form

Reference 132

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T08:49:42.711489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:f86febd682bd1513e1d372c94a39a8dee6b73296ac0551055d01a2868b1ca9ab