Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T21:36:02.340129Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2511.15053.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T21:36:02.340129Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 202f774e-7096-4356-8161-fbac868d4dc5 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Multiagent deep reinforcement learning for large-scale traffic signal control,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d13c1a4a-3f71-417a-a326-9f89467e8d72 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Large-Scale traffic signal control using a novel multiagent reinforcement learning,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c33e2ac0-400b-4c72-92b7-d082f667f0a2 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Applications in traffic signal control: a distributed policy gradient decomposition algorithm,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e140e33f-ce4c-4ac2-952d-cf8a50967564 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Deep reinforcement learning for joint channel selection and power control in D2D networks,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b90cb7-8e76-446c-9042-fbcb52c5ffac · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Power allocation in multiuser cellular networks: Deep reinforcement learning approaches,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6551d4cf-8a2c-49a9-bedf-daf6412588db · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Reinforcement learning based recommender systems: a survey,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f42c7288-b59c-461e-9a3a-cdf6b359ac81 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A survey on reinforcement learning for recommender systems,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbdec37a-c4dd-44c0-bd48-4230259fe7d2 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed reinforcement learning algorithm for dynamic economic dispatch with unknown generation cost functions,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b88c68d-1cd0-41ee-ad15-76f428da4d4e · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies DistributedQ-learning-based online optimization algorithm for unit commitment and dispatch in smart grid,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f5a3812-2028-42cd-a30c-b04df87ab78c · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed Q-learning algorithm for dynamic resource allocation with unknown objective functions and appli- cation to microgrid,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0196251-46f1-488b-ab98-d0effe1a17b4 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a90f67-1924-48f4-870c-66dba815b9eb · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Con strained reinforcement learning has zero duality gap,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76579216-ff66-4fe1-8d2f-1364ca164127 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Safe policies for reinforcement learning via primal-dual meth- ods,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 507e94b4-2379-41f4-885b-6244e5451ab2 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies State augmented constrained reinforcement learning: overcoming the limita- tions of learning with rewards,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3cbf48b-7b5b-4b1a-9a0c-5931cbdb46fb · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Probabilistic constraint for safety-critical reinforcement learning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a944c460-4235-4e10-9049-9b4bea9313bc · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A review of safe reinforcement learning: methods, theories, and applications,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27dbf847-1a01-46c4-991f-c00a01771f26 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Provably efficient safe exploration via primal-dual policy optimization,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef70a023-d5fe-4faf-88a2-7ef0bb92289c · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A dual approach to constrained markov decision processes with entropy regularization,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f90b93-a4ed-4bdc-ab33-7b00d208e04e · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Policy-based primal-dual methods for convex constrained markov decision processes,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34992d50-10d7-4009-a6e8-284b92ae6abb · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cc5165f-3152-4e34-96ae-8dc5aaf7b2c0 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Scalable primal- dual actor-critic method for safe multi-agent RL with general utilities,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11f0d457-b2ef-49fc-bac6-1c8e13c3703c · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Scalable reinforcement learning of localized policies for multi-agent networked systems,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51f2b191-18ba-4928-9712-df6d1be8f032 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Solving a class of non-convex min-max games using iterative first order methods,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 471f8719-292f-41e7-a22c-eec62f550c55 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Policy gra- dient methods for reinforcement learning with function approximation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07ed2a0c-8741-4788-bb33-d73021d81a1b · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed optimization over time-varying directed graphs,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd4d33d2-afbe-4190-931c-843e0bbb0e49 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Zeroth-Order Policy Gradient for Reinforcement Learning from Human Feedback without Reward Inference
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ab3935-7daa-4efc-a47d-36a7d9d5a556 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies ϕ-update: a class of policy update methods with policy convergence guarantee,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d16c18a-1f07-4271-b429-0b7fddc57f83 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed stochastic subgra- dient projection algorithms for convex optimization,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f22a796-18e1-43b1-9b1f-bb53375b693e · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Yeh,Real Analysis: Theory of Measure and Integration
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a694b7-2513-4fc3-8b3a-58335de94770 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Global convergence of policy gradient methods to (almost) locally optimal policies,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e7dcbf5-7854-4ae8-a71f-9cafcc0e60c2 · outbound
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Given that the proof procedure is strikingly analogous to that of Case (i), it is omitted for brevity.2 G
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.