Pith. sign in

Paper Citation Record · LEDGER

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions

As of 17 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 0 inbound Pith citation observations for arXiv:2608.06545.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06545 v1

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:39:15.637632Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 300 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2ccdd33f-3f04-4beb-a6ae-793c30b353ad · outbound

This paper cites Towards Tight Bounds on the Sample Complexity of Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Towards Tight Bounds on the Sample Complexity of Average-Reward

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.203436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.203436Z digest=sha256:58c447e136af6ca92182031e015044b28e559250d6f3c3cbf63f9368e8f94966

Observation f2cc61b9-093d-4e5a-a0a2-8c15624c0184 · outbound

This paper cites Foundations and Trends in Machine Learning , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Foundations and Trends in Machine Learning , volume =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.207525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.207525Z digest=sha256:dcfc874a8484de6cae3a921c84009f12360eb65fe25242e1dc71b01d886355e7

Observation 8ea6c8a4-068c-48d2-9b23-c30a2867e586 · outbound

This paper cites Operations Research , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Operations Research , year =

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-15T14:39:17.218618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.211143Z digest=sha256:c151803d6fe33f8d71cd9aca372c685ea48b47e827724b0cf7426ce04a1e4215

Observation 18cd24de-c12a-4c63-8b12-ff2875b2223e · outbound

This paper cites Near Sample-Optimal Reduction-Based Policy Learning for Average Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near Sample-Optimal Reduction-Based Policy Learning for Average Reward

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.215035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.215035Z digest=sha256:13b2ff8d7e6abbbb56b697ef68e7801d86ed09cef9524f1d337f5b1be84d61a8

Observation 308a30ac-f493-43a3-bd7b-cc343280fed5 · outbound

This paper cites Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.218684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.218684Z digest=sha256:a97610f2c6543c3dd7a1793e0ef11f03cd7cb7f49e94097528c60455937fb50f

Observation 07ceabbc-b080-41b6-adf4-74793e9a4c3b · outbound

This paper cites and Tewari, Ambuj , booktitle =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions and Tewari, Ambuj , booktitle =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.222526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.222526Z digest=sha256:2db61ac3b4c436788558c5a4e0f963381e0a9d92cc68552baa4c2a323f625b53

Observation bf3c2cea-2013-4688-bb5a-c88449183c95 · outbound

This paper cites Proceedings of the 35th International Conference on Machine Learning , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 35th International Conference on Machine Learning , pages =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.225630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.225630Z digest=sha256:288ed1d1a8c2951645c15292a76f83300879d57818975150c462263f81ca8e33

Observation a9581ca9-bca9-44e1-b18c-66313c1fc4a2 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 42nd International Conference on Machine Learning , pages =

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.229174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.229174Z digest=sha256:29207be01c11598790e17de93084ddf8713fbb7f9e73fe885e2cb38cf8535fe2

Observation d541f5cd-ef54-46b1-8f06-ffb1343a118f · outbound

This paper cites Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.232280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.232280Z digest=sha256:b50bcce9d23ee06e8ef11a251e1b4e125b779656fcd552595f20a513f33d4b8c

Observation e4a6617e-4655-40c1-86f7-3fdbac929a7e · outbound

This paper cites High-Dimensional Statistics: A Non-Asymptotic Viewpoint , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions High-Dimensional Statistics: A Non-Asymptotic Viewpoint , year =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.235726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.235726Z digest=sha256:4e98aaf5945108f69442eb99a6c31d201fe366acb1cba253d17c5f21405fc735

Observation 70139650-2f18-4ab0-b2de-5dfa75baba63 · outbound

This paper cites arXiv preprint arXiv:2603.00945 , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions arXiv preprint arXiv:2603.00945 , year =

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-15T14:39:17.144034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.238646Z digest=sha256:8eec44251fd42ff130148797cd7e1bb4ca199a75b2b0996b32423767003ed23c

Observation 9f37f0ef-f166-4472-8cb7-778a0fd55dfe · outbound

This paper cites Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning , year =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.241863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.241863Z digest=sha256:2d553909dcff02d7695bbec03be59c5dfb0792275527ccc1af880d4564cc7725

Observation 3e7d40ca-995e-44ca-9dc9-ed77e017b67b · outbound

This paper cites Efficiently Solving.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Efficiently Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.244721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.244721Z digest=sha256:fd330dd026ba1058cf4ce7d2d5b48d04f15e984b73086c06a928b81d10722526

Observation ae41af26-e4af-445c-aae9-695e441f1fc5 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Twelfth International Conference on Learning Representations , year =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.247625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.247625Z digest=sha256:c603174a091c0e5832c5a5f4445e968386e21e2553228cc83364721e3326a624

Observation 76f441f9-afbb-4805-87b0-bc8a917a1e15 · outbound

This paper cites The Plugin Approach for Average-Reward and Discounted.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Plugin Approach for Average-Reward and Discounted

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.250801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.250801Z digest=sha256:500bb46fffbe37a710e5c1c014725a17b2ba5aa7c52942ae6b5b8b49708496e9

Observation 88f0423b-98b0-4821-8345-22477419d669 · outbound

This paper cites Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.253403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.253403Z digest=sha256:591e4ce555b3d07658fa49cddf6e11d6f756ad468104bb95c1562aa42c879d61

Observation effe7f4f-ba8d-490d-9f9b-38e263cabdf7 · outbound

This paper cites Sharper Model-Free Reinforcement Learning for Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sharper Model-Free Reinforcement Learning for Average-Reward

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.256314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.256314Z digest=sha256:e429bf03a65d9231cbf60b5323f1009b669d93ed6710883064f07324256b3815

Observation bc7cecdf-4059-422f-8f9e-50a6354b90d3 · outbound

This paper cites Journal of Machine Learning Research , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Machine Learning Research , volume =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.259013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.259013Z digest=sha256:ff60c032eb21183fed1dab379b5de9bd747323f4b38c58083cb59e216e2604d8

Observation 5866a8ca-c052-40bf-bc7f-c07cdd22a80d · outbound

This paper cites Model-free Reinforcement Learning in Infinite-horizon Average-reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Model-free Reinforcement Learning in Infinite-horizon Average-reward

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.261581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.261581Z digest=sha256:d9b898a8fc3b4d25b715faceb17c7e6f7ee4d37ce67468b93953b9c60b6f5697

Observation bda1fb27-dbc1-44e4-8275-9035f35c413e · outbound

This paper cites Learning Infinite-horizon Average-reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Learning Infinite-horizon Average-reward

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.264641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.264641Z digest=sha256:960550e080c20e6d9c0b9b86e65239e6742a897aef297bdca02dbe6cde9df13e

Observation 52a1e67a-ab1f-4cdf-8bfd-44fdaeb2fcea · outbound

This paper cites Efficient.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Efficient

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.267560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.267560Z digest=sha256:9b20e0fd0afef891d96f3ff9476cc6002d29cb62e933d26cde6507db2ad0fa2a

Observation 060f7f15-ac12-4f1e-a64a-c3b93494a9ab · outbound

This paper cites Distributionally Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.270443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.270443Z digest=sha256:06cd061af4b3b3bc33751ac2246df04553449828c049685cfce4536a84c1f759

Observation bb522753-d61b-4193-9151-66965ad1ffe4 · outbound

This paper cites The International Journal of Robotics Research , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The International Journal of Robotics Research , volume =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.273226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.273226Z digest=sha256:c979438b737e2cbd0d56924fa8017e62e998eede9d56a9e456fc359f47090a64

Observation df27e9a6-4869-4efa-af6f-0ab6da554678 · outbound

This paper cites Nature , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Nature , volume =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.276155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.276155Z digest=sha256:9cb71c803d91b1680b4474baf1742d98330db9daea94be1f587d13d8a81ddd10

Observation 7f30be99-0561-4d7a-b0ef-3ab61d57ac2c · outbound

This paper cites Mastering the Game of.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mastering the Game of

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.279523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.279523Z digest=sha256:307538b34d5a77dff96d2eb5c8ac53e2f39197ee956e2b7335a6172b64ad2b0c

Observation cdd71e41-3f14-437e-a229-b55c90a318e1 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.282444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.282444Z digest=sha256:27628107617740d195715ad6cba9b51ff552365274b6205a89bd7d8ec8ee9361

Observation 2b7b49e0-2026-41da-9bd1-bff29f14ac5c · outbound

This paper cites 2020 , organization =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions 2020 , organization =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.285478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.285478Z digest=sha256:c89366db30e69d7dbe56d068f119fa6fee31d8fbde2ee8c960c4991746346595

Observation d48defea-6d63-4673-a321-6bd36d16a06d · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.288546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.288546Z digest=sha256:06d99e7416e1752c3cda8427b1ee58c59d53f309f93fb6938034efc5d7ff761c

Observation f24d8ef9-b189-4daa-b706-59e90a279f67 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.291628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.291628Z digest=sha256:de9720ef47e5e2a7027c8d49d8268fc54f222a4a49c17999524f3039c75dcb0f

Observation ab448d6c-4b6e-4cca-84fc-118f4a32ca85 · outbound

This paper cites COLT 2009 - The 22nd Conference on Learning Theory , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions COLT 2009 - The 22nd Conference on Learning Theory , year =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.294253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.294253Z digest=sha256:ad4f1f9cddaef5e069c14b84f2f073697499cb0f13816aaf980af869e8a1b795

Observation caf66f2d-f67c-4a94-88f2-c20bd87974ba · outbound

This paper cites Data-Driven Distributionally Robust Optimization Using the.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Data-Driven Distributionally Robust Optimization Using the

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.297291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.297291Z digest=sha256:0195b9cd1439df4b65c15a7f34433c41e2c886637e0c3838d1a228117db0d1bb

Observation dadd3c77-37ac-4010-a482-bbc94ff2f375 · outbound

This paper cites Distributionally robust convex optimization , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust convex optimization , year =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.300073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.300073Z digest=sha256:7997db569e8042ed6eace020b151286cacb7973244089bea4231b378eb7c72d5

Observation 8d85ffe9-de8a-4cd5-99c2-9f2c346e662b · outbound

This paper cites Distributionally robust optimization and its tractable approximations , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust optimization and its tractable approximations , year =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.303083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.303083Z digest=sha256:2388c24153a492aae0fa4b9b030888496a1ab6b8c9cc08560717c5c8ab1186e6

Observation 5066e41a-5384-44dc-aa61-aa46226213b7 · outbound

This paper cites Learning models with uniform performance via distributionally robust optimization , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Learning models with uniform performance via distributionally robust optimization , year =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.306435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.306435Z digest=sha256:756993a63d3cc8597d3444c7348c0f3eaa1c7b9c04ab8adebfa7f056ec23e2fd

Observation ef0a72be-43a9-46dd-b59e-e6da2f75aab8 · outbound

This paper cites Robust Average-Reward.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust Average-Reward

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.309479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.309479Z digest=sha256:3a2774869106536229b904c6b3aa924f0e268a8b66b4b6e44dde97a29803d635

Observation 13d7ae5b-197e-4f50-99e9-e2f79fbf928d · outbound

This paper cites and Prater-Bennette, Ashley and Zou, Shaofeng , booktitle =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions and Prater-Bennette, Ashley and Zou, Shaofeng , booktitle =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.312817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.312817Z digest=sha256:c4e48935ecaa6b3fb0bdcdad2a98f5f01a517152e7cb8bd8ca1af0be13033e44

Observation c85eb2ac-fcb5-4b2a-969e-a4e561107ce5 · outbound

This paper cites Toward Theoretical Understandings of Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Toward Theoretical Understandings of Robust

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.315761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.315761Z digest=sha256:c47f3e7fdc02dc20fc120c21ef99c17f60c6fc2719bc2e5fc7e7cdf943539f15

Observation c9a23339-85dc-413f-8167-7dc904fef751 · outbound

This paper cites Sample Complexity of Variance-Reduced Distributionally Robust.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Variance-Reduced Distributionally Robust

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.318707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.318707Z digest=sha256:4f872c8f7e86056aaefefabb496b7e87335653333aa7c23801cb68ae7f65a931

Observation 5b71829d-561e-49c5-9191-50ddc86959cf · outbound

This paper cites Near-Optimal Distributionally Robust Reinforcement Learning with General.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Distributionally Robust Reinforcement Learning with General

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.322481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.322481Z digest=sha256:54642b9061458524b35d269cf53cc914ab2b927d13dbaa943c812ed0d9ebf2c0

Observation 51e9ba48-0ed9-4ecc-8bc8-9d3e96119b5e · outbound

This paper cites Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity , year =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity , year =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.325515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.325515Z digest=sha256:a0828eb817822cd782a081412a5f9ec826f537312ed56f68ecc6ce68db11ce5a

Observation 82caf014-8a07-4f00-a1b6-b4f5deb0e0e0 · outbound

This paper cites Sample Complexity of Offline Distributionally Robust Linear.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Sample Complexity of Offline Distributionally Robust Linear

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.328595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.328595Z digest=sha256:d4dd3cb91d123292a1f2aed5247e2ca18fc11319af535c58a354d081110ca25c

Observation 0277484c-a2c1-4750-8320-96faf0343c2f · outbound

This paper cites 1994 , publisher =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions 1994 , publisher =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.331697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.331697Z digest=sha256:f4a88d1dee257618732b5239490acd714a94e1c2d06dfb45ffe3f603e9ea4851

Observation 0acdd3e8-5b79-46af-849f-79700514230d · outbound

This paper cites Tsybakov , publisher =.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Tsybakov , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.334697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.334697Z digest=sha256:70248fb0ea300b0aeaca28de35983697963d25a648cb95d08caf1e12be39e2da

Observation f9b615a3-36c1-47b2-a5fc-34953045dd77 · outbound

This paper cites Machine learning , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Machine learning , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.338841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.338841Z digest=sha256:ac123b5aa656c8d044e3f12a678dee080c567b5ae32d8be51cbda3881273c68b

Observation d42131d3-33cc-4460-a376-4f7a909dfe1e · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.341917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.341917Z digest=sha256:ca481e46899f87e261c646701229a131830488a2885c736cff3b278ebd4aa50b

Observation 219b76fe-3d4c-467a-afbe-9f27746d6e13 · outbound

This paper cites Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.344469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.344469Z digest=sha256:4cdd63904352acdc079dfd457e96b1e5ec52a3ef76634a9be94e36309bf22d26

Observation d613f426-39b7-478e-b447-630fadb22f2e · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Forty-second International Conference on Machine Learning , year=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.347703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.347703Z digest=sha256:54996a290a0c42260b850d6fc69e3ebed12fec1c240fa84db310473ea04fd139

Observation 9babee25-36ec-45a5-8cd0-f2e18149d86e · outbound

This paper cites The blessing of heterogeneity in federated.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The blessing of heterogeneity in federated

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.350738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.350738Z digest=sha256:aa246bc81a2bc6411345894c41ccccbf9930fc9bb2df41afec6fd8b99b46f522

Observation 3d5003de-2530-4f7f-a5e2-59e8f1b38eed · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.353842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.353842Z digest=sha256:dc3b1f8274c5a1c7d5557ec6bcef42c47c5771f595d5cbf6fd10de22b8a2d62c

Observation 3b7dc32c-427a-4975-b155-89302f03eebd · outbound

This paper cites Thirty-seventh Conference on Neural Information Processing Systems , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Thirty-seventh Conference on Neural Information Processing Systems , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.356422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.356422Z digest=sha256:224c507628ab3b707652c7fc1a983711a3ffbf0a79c0d495fec0c987f4bf10cc

Observation d03d0526-5a80-4b81-b762-847536a57436 · outbound

This paper cites Operations research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Operations research , volume=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.359852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.359852Z digest=sha256:8a1795b6bb83823b0371078eec93f154d679b818752cffef6f4e756b0649a2a2

Observation a3b7a8dd-08e4-44e2-9442-3a14b31ff866 · outbound

This paper cites Robust control of.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust control of

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.486678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.486678Z digest=sha256:8a06e85067b772ce76f13fe789a3543e218fbe319a6d18908e3fdebf6751f6fd

Observation cdfc98d3-e4d7-4f31-bdef-0eeadf2f775e · outbound

This paper cites $Q$-learning with Logarithmic Regret.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions $Q$-learning with Logarithmic Regret

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:17.048249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.493614Z digest=sha256:54cff01636a73a461f02573969084240634f6bd488fefbc17cdeb8e5ac0520b4

Observation 791b6d3f-07a0-4b23-b368-86e548e84c4f · outbound

This paper cites Near-Optimal Provable Uniform Convergence in Offline Policy Evaluation for Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Near-Optimal Provable Uniform Convergence in Offline Policy Evaluation for Reinforcement Learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.496949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.496949Z digest=sha256:fce10482cadbc92aaf1a81439bfd2bb42c479368b78642c833cb606adfd6c8d5

Observation 78457264-6cba-4039-9fd9-2f848b71ca06 · outbound

This paper cites Advances in neural information processing systems , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , pages=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.500102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.500102Z digest=sha256:91db812b71c4a501c82ff3e783004f0efe1ca43425dd377756de41a4ee27aba3

Observation 4d4d9a1c-2280-4df3-8905-23cfc31a8ac3 · outbound

This paper cites Complete Dictionary Learning via $\ell_p$-norm Maximization.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Complete Dictionary Learning via $\ell_p$-norm Maximization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.503037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.503037Z digest=sha256:d339fa7e22fc5e2f06baba1c46bf9d80e163e17627cab97f052c30209c8e21bf

Observation 4c2ba3de-423a-48a3-993f-8e7b0fc6da5f · outbound

This paper cites Journal of Applied Probability , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Applied Probability , volume=

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.506776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.506776Z digest=sha256:f73189d28ac6355349a38671c1269209731d44194da5dbdf22bdbbee35709dab

Observation d88536b2-f16f-4da6-955f-bbe6e9d0198e · outbound

This paper cites Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Non-asymptotic Convergence Analysis of Two Time-scale (Natural) Actor-Critic Algorithms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.509987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.509987Z digest=sha256:45261288e8d46d4ac30c0a3d41dac1bd806aafe9264c40def7e42b475592b0c0

Observation b285e844-8231-4ac4-b382-a11c23509a4c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.513098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.513098Z digest=sha256:1ea3b66f06d504d35d5f55fa755f49e74551fd8f681b5b5eadea0418a98a075e

Observation a2746f34-f148-4e9d-b074-5d3f586fae8e · outbound

This paper cites On the Global Convergence Rates of Softmax Policy Gradient Methods.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions On the Global Convergence Rates of Softmax Policy Gradient Methods

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.516180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.516180Z digest=sha256:deb0d6c4771565519398d10832e914a586bfa4d76b37ed0880b87f239c8ab455

Observation 9bfa37da-e391-4468-aad1-830e70ba60a6 · outbound

This paper cites ICML , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions ICML , volume=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.519221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.519221Z digest=sha256:6eb69d977d274198fd7bd8b1cd3bf0ea26dceb98bdee9d5e982d655d3e35659c

Observation 24b203a2-f37d-4dad-9454-be3634036b70 · outbound

This paper cites Optimality and approximation with policy gradient methods in.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Optimality and approximation with policy gradient methods in

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.521951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.521951Z digest=sha256:3a34b67d234bef6bb0e3f962e93d12776a8b8290a868da1efd87b9e99da36502

Observation 4619700e-2526-4d70-b16f-2937434744ee · outbound

This paper cites Advances in neural information processing systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , volume=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.524694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.524694Z digest=sha256:b31149d57add3a3d0478e8f2d13c97bf0a9907de97badaaf5b3851b9fd01d8ae

Observation 6acaff08-1837-4d96-8ec8-579d4069fbd3 · outbound

This paper cites Advances in neural information processing systems , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , pages=

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.527449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.527449Z digest=sha256:a267da14da2b947bbb4ded3602ed24d4c7be9ee9837335b746a7944589a9eadc

Observation bf8226aa-05b0-4c66-bf10-6907e51e48cc · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.530633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.530633Z digest=sha256:61b2fa029e8cb340d3ea4842deec605b5a8c5d186bd46e4af7c30fea8f3c560a

Observation bb02d565-1b38-4e88-ba39-9677c0ba4def · outbound

This paper cites Nearly Minimax Optimal Regret for Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Nearly Minimax Optimal Regret for Learning Infinite-horizon Average-reward MDPs with Linear Function Approximation

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.986137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.533537Z digest=sha256:090dbe271930525d37938c8d2ad5817326bf757a8d9a2ab4905af7181b494d21

Observation 41ab3e61-0120-460f-adce-2841f65c14c5 · outbound

This paper cites Probability Theory and Related Fields , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Probability Theory and Related Fields , volume=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.536769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.536769Z digest=sha256:f9f50d0205314ada6823b7d60c16f99e7bed5a6c4e1b2c2d400dd0081d8f724c

Observation ad91d121-3c2d-4722-8a4f-b3625fa945fd · outbound

This paper cites Proceedings of the 27th international conference on international conference on machine learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Proceedings of the 27th international conference on international conference on machine learning , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.540035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.540035Z digest=sha256:86d6c2a058798f52eb68c95a4e1496797c29960d314c6bd3e242ef70bac7a13e

Observation 61a978c1-dc40-42d6-8a93-e2f6ea1fb7fa · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Learning Representations (ICLR) , year=

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.543051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.543051Z digest=sha256:4b2d7f3bf6513a5f9bc109d70f4ce274e7b00c860c215d00cdb504ac87df8ff3

Observation 5a63e9ef-e709-4519-a2f2-b938816c1d4e · outbound

This paper cites Theoretical Linear Convergence of Unfolded ISTA and its Practical Weights and Thresholds.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Theoretical Linear Convergence of Unfolded ISTA and its Practical Weights and Thresholds

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.545887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.545887Z digest=sha256:e940d02707b47afd7443e46087c670ec3781f2b3ff9fcaa4cbec2019b60cc8b2

Observation 835631c6-af5a-4d43-8270-834dab3ff1ba · outbound

This paper cites Ada-LISTA: Learned Solvers Adaptive to Varying Models.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Ada-LISTA: Learned Solvers Adaptive to Varying Models

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.960039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.548922Z digest=sha256:0a7dcbbb49391537401664d12985ca3cce0ef64a0261b4871cbe3ede13cdfecb

Observation caac32fe-da59-4a36-a272-c0efd64b4a3c · outbound

This paper cites Understanding Trainable Sparse Coding via Matrix Factorization.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Understanding Trainable Sparse Coding via Matrix Factorization

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.552169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.552169Z digest=sha256:d8cb92b33dc26f0879001f9cd2a63ce503d0a4d9bf0d1da90e17fc45e6749139

Observation 14528f54-1f2e-4b30-ab9b-7c9c74523b67 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Artificial Intelligence and Statistics , pages=

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.555278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.555278Z digest=sha256:47529c9064c538c53e3e7ea9df61b243582e8f27241858d3994cb1edb059bc22

Observation bc2f6af1-29a9-45fd-b665-3ca9bf0dc9fe · outbound

This paper cites Mathematics of Operations Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematics of Operations Research , volume=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.558053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.558053Z digest=sha256:96422c50ded0a7668546e6397b3d34b2618791cc6ba4d91a21b96b8fa80257d5

Observation 3dde4795-b613-4951-b572-7d0e891208ea · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in Neural Information Processing Systems , volume=

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.560606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.560606Z digest=sha256:a288c6d411321416d52e5ba0b4e3865ace4b6df19d099673fd55d9a4e634e81a

Observation 0e6eac17-e3d8-47c5-96cb-ecc2c259c858 · outbound

This paper cites Twice regularized.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Twice regularized

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.563461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.563461Z digest=sha256:1eb87d1cee6499377080029ac41492f939ce6fe830d7123c8014d8dd96f997cf

Observation 5e06f2a4-1d0e-41aa-b21b-c3ef8b779bcb · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.566048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.566048Z digest=sha256:114e67d3a71b84abd884534b6f7f71ed0232a354ec4a0810db9dbbe718e17b95

Observation 1ce8b195-a019-46e1-a093-05fdf0f79678 · outbound

This paper cites A Review of Off-Policy Evaluation in Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions A Review of Off-Policy Evaluation in Reinforcement Learning

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.568605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.568605Z digest=sha256:e3f38e15443b477d31684d303061989298ec3e43e40f6015086c3b75200e9929

Observation 73c7a7ad-9d56-440f-a888-20da053818af · outbound

This paper cites International Conference on Machine Learning , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions International Conference on Machine Learning , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.571404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.571404Z digest=sha256:48eca39456fc3cc4903eae3ccf2341b7f0483cfc70d636759d093ec44698e4ac

Observation 1fd2fac0-ba16-429b-aad4-5d1ca99d20bc · outbound

This paper cites Distributionally Robust Optimization: A Review.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust Optimization: A Review

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.574358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.574358Z digest=sha256:37f4dfb15bbc64f260ae7d597b6f9d6e0e55816c7375334776d863eaf02d73b3

Observation 95b26c2f-0f01-4095-a4a2-0ece99f3e556 · outbound

This paper cites Finite-sample guarantees for.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Finite-sample guarantees for

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.578729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.578729Z digest=sha256:5c958fc9d9b745f5fe07e63e9ac3ea88ebc0d904a5c80bcc9806c75a882eaa8e

Observation 11359dc3-2b4b-499a-9887-c87535093a98 · outbound

This paper cites Certifying Model Accuracy under Distribution Shifts.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Certifying Model Accuracy under Distribution Shifts

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.581826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.581826Z digest=sha256:bee5faa403c09c03debde5e50f3c98ea9d78f60cb36710ee22a42edc2fab7aa9

Observation a76b1b01-4c57-423d-a0bf-bc2a6f1facfb · outbound

This paper cites Advances in neural information processing systems , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Advances in neural information processing systems , volume=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.585147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.585147Z digest=sha256:1398f14d2d21c0eae9699d1a361560a717038792c7a93e4570f756b2f14a5548

Observation 3fd88426-9fe4-4008-b4d0-9a108e0bdd80 · outbound

This paper cites Settling the Sample Complexity of Model-Based Offline Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Settling the Sample Complexity of Model-Based Offline Reinforcement Learning

Reference 84

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T14:39:16.893053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:39:15.588609Z digest=sha256:cc6367b2bab55eeeeb074540d807666f7aaea99993f91cbc32d7ef83c033caf5

Observation d72da13f-f3f1-4232-974a-19cff24db1e6 · outbound

This paper cites Available at Optimization Online , pages=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Available at Optimization Online , pages=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.591949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.591949Z digest=sha256:e836e532e0c753d2070b7f065cfeaf947cb80e7ef1af307d6cf62aede22c1648

Observation 81a5dc8b-44de-43f4-97e9-b1e3e9a9078a · outbound

This paper cites Pessimistic.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Pessimistic

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.594947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.594947Z digest=sha256:3a2b74158080334526c2c8c82c32a08939a78596688adb3520b668e74eabc2cd

Observation 1ca62b82-0e54-4ad5-94c9-e6599cd5f221 · outbound

This paper cites The Bell system technical journal , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Bell system technical journal , volume=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.597974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.597974Z digest=sha256:dc3ab4d09c71fdc3bb20a6a6917d15f0eaca967066fcb8d3e556e9f9d1ab8f50

Observation 3bf77a69-8377-415e-9085-5610c6dd79f5 · outbound

This paper cites The mathematics of data , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The mathematics of data , volume=

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.600564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.600564Z digest=sha256:be25ef25cd447ce9870068d8fa3bb0f6330b04219ecc637a47d2aee9195705d9

Observation 0e2948c7-0976-4bd5-a85a-bce669fd91ee · outbound

This paper cites The Journal of finance , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions The Journal of finance , volume=

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.603762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.603762Z digest=sha256:28844dae59e5f4feba19544f22d4ed0d52072ab042782edccad587b6a4bc095e

Observation a06183ee-4d5e-455b-a185-f794e2047b6b · outbound

This paper cites , author=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions , author=

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.606758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.606758Z digest=sha256:13729face493478e9142034869c26a786422b7cb190d7daf75571fd7ea74f50f

Observation 47828712-199e-4990-a716-666c81d0da09 · outbound

This paper cites Mathematical Programming , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematical Programming , volume=

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.609707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.609707Z digest=sha256:c0644f72da74120ddd094bc8a731792b86162126fe7858f2e700d0e49c743981

Observation df1e378e-fe72-4bd8-b904-22dd29d668a4 · outbound

This paper cites Mathematics of Operations Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Mathematics of Operations Research , volume=

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.613255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.613255Z digest=sha256:d3f3761a1924a9130ae11dd7c78fc1234f5f73898e4def778781f1e964c45222

Observation b86932a2-17df-4343-9aa1-13f973b37687 · outbound

This paper cites Robust control of uncertain.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Robust control of uncertain

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.616496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.616496Z digest=sha256:5a99f2982cf39bfe009adef3dd7dcb7c9642ae288b8f72141857713262e6a06c

Observation a521443e-c233-48e4-876b-650f335feb2a · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.619844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.619844Z digest=sha256:5a8b669da823389331304ffe9db9f3c972e23677bcd3ecddf5bf0233b86f4151

Observation d86bc39b-ddcb-44e3-8a07-023b3e422c4d · outbound

This paper cites INFORMS Journal on Computing , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions INFORMS Journal on Computing , volume=

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.622784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.622784Z digest=sha256:621d952e404fba074149f3aa959e88cfcfd11c2274975e33d875272d37f3457c

Observation c290e883-4b9c-42ec-b04f-c5ed1283589c · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.625757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.625757Z digest=sha256:d31d8276b506f6a6c7c7bfda6a1d596fb73ee77a9e7c8642ad411b688de99d9a

Observation 29f6aba9-0e16-4f56-af81-020f27f8653c · outbound

This paper cites Distributionally Robust Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributionally Robust Reinforcement Learning

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.628743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.628743Z digest=sha256:877c8cc806407a9befe90a6675e9e35a9482c0047beb608cc922f0d81684b647

Observation e8b0c55a-7f0a-445a-a4ca-a48965dd18c3 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Journal of Machine Learning Research , volume=

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.631794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.631794Z digest=sha256:282acb9a834bd08cdf52a9edfe26214b7ac36b73dfc3aa21088eeef546dc47c7

Observation a6bed617-be62-448b-b66d-d30b9ee729a8 · outbound

This paper cites an unresolved cited work.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Unresolved cited work

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.634562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.634562Z digest=sha256:a941cb3f8e53023773ea30fee047121965ffab1b0082f3f8e387eeda1a9a4977

Observation 88f1c433-0416-48a4-bdf1-f5f8f2a8cef7 · outbound

This paper cites Distributional Robustness and Regularization in Reinforcement Learning.

Robust Average-Reward Markov Decision Processes: Minimax-Optimal Learning via Plug-in Reductions Distributional Robustness and Regularization in Reinforcement Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T14:39:15.637632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:39:15.637632Z digest=sha256:432603f45e574b1ffdaa0ba952b1f7daec2444176cf91cef2f7f57035d1e3e31

Pith citing papers

No inbound Pith citation observations are available.