Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:10:22.389652Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2412.19873.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:10:22.389652Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 915506fb-b58e-4222-b634-75fc90fcb3ae · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Minimax-Optimal Multi-Agent RL in Markov Games With a Generative Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f938e5a-86e0-4651-8374-64b4a7743f45 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Markov games as a framework for multi-age nt reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d2895eaa-edec-4d2a-8aaa-4105d8e89c17 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning When Can We Learn General-Sum Markov Games with a Large Number of Players Sample-Efficiently?
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fc7b9a9-a76c-4ba5-823f-38be90703761 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Robust Markov Decision Processes without Model Estimation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3ce7b97-a151-420a-ba97-31cf6157b70f · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning $O(T^{-1})$ Convergence of Optimistic-Follow-the-Regularized-Leader in Two-Player Zero-Sum Markov Games
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2a7760c-9654-4f32-9bd4-e2096752ddfa · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning SustainBench: Benchmarks for Monitoring the Sustainable Development Goals with Machine Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ebef0ff-5a5c-4bd3-8399-964dae04b62f · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation addabe88-3258-4706-b570-cb1396784e6a · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning URL https://onlinelibrary.wiley.com/doi/abs/10.1002/j.1538-7305.1952.tb01393.x
Reference 1952
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 39163fcf-0cbe-496c-a08c-d6efa405b086 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Breaking the Curse of Multiagency in Robust Multi-Agent Reinforcement Learning
Reference 1953
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02150360-0090-4cf1-981c-c15e25cc39c5 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning OpenSpiel: A Framework for Reinforcement Learning in Games
Reference 1998
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ea42e5f-01e7-4d3b-8e87-e50599d88246 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Feature-Based Q-Learning for Two-Player Stochastic Games
Reference 2005
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddeba19e-c782-45bc-ad07-ea24804a82ec · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68fb79ea-9549-4541-8865-d5c546ea554d · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning V-Learning -- A Simple, Efficient, Decentralized Algorithm for Multiagent RL
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cbd4675-0fb1-45e1-942f-2c23bf85c093 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning DeepRacer: Educational Autonomous Racing Platform for Experimentation with Sim2Real Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573bbe92-7adc-4381-a30f-f0deaafe9f67 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning SMART-LLM: Smart Multi-Agent Robot Task Planning using Large Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b44bf1-66ce-4e3c-8873-a7c3b9a4e011 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning Fast bell man updates for robust mdps
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 352ad0fb-5974-488d-9a31-d218b7db4354 · outbound
Minimax-Optimal Multi-Agent Robust Reinforcement Learning What is the Solution for State-Adversarial Multi-Agent Reinforcement Learning?
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.