Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:16.554175Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2505.24113.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:16.554175Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f776eeba-bc9d-4697-8e94-9e8b33660b12 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 177ae6f0-28fd-4767-81c8-c2de3da67f73 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Distributed reinforcement learning algorithm for dynamic economic dispatch with unknown generation cost functions,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c79187e-5f15-46d6-ae2b-f0c572a8aa00 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Distributed Q-learning algorithm for dynamic resource allocation with unknown objective functions and ap- plication to microgrid,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 124b19d0-8fb3-461e-a3c6-092c2fe33ac4 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Multi-agent deep reinforcement learning for large-scale traffic signal control,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e346283-6c11-47df-84f6-9088b10b160a · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Large-Scale traffic signal control using a novel multiagent reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 81656743-995e-421b-9623-141de60b2656 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning A multi-channel transmission schedule for remote state estimation under DoS attacks,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fb7678ca-e2ce-4a61-b5fa-3fa4f5aec444 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Distributed reinforcement learning for cyber-physical system with multiple remote state estimation under DoS attacker,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acece6bf-4ca5-4592-ab63-789a746ad2c0 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Multi-agent deep reinforcement learning for dynamic power allocation in wireless networks,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b150f2d-2476-40ad-9c27-d60063479963 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Deep reinforcement learning for user association and resource allocation in heterogeneous cellular networks,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63c2aa6c-a616-4b83-8a28-24d140a4c9b4 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Intellilight: A reinforcement learning approach for intelligent traffic light control,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40a04de8-ddbb-4eda-84ac-7ad3fc482955 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Multi-agent reinforcement learning: Independent vs. coopera- tive agents,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b7b1cbf-c919-4ebe-b79d-1eb9ec6c88c9 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a4adffe-50e4-426c-bca7-870c67feafc7 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning QPLEX: Duplex dueling multi-agent Q-Learning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49141aae-5ed1-4523-ace7-eccadf984a04 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Counterfactual multi-agent policy gradients,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18c2cc5b-0075-45e4-9a09-ef2c9bdcf37f · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Off-Policy Multi-Agent Decomposed Policy Gradients
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dbe82342-d029-4363-9ff8-28d99790f797 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Fully decentralized multi-agent reinforcement learning with networked agents,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a51f0f31-88cd-4519-a326-0f3338064334 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Distributed actor-critic algorithms for multiagent reinforcement learning over directed graphs,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b2ccda6-253a-484b-9872-73a39ac75c81 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Neural policy gradient methods: Global optimality and rates of convergence,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14762a8d-a7ae-406e-aa1c-8f7276f78ecb · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Finite-time analysis of dis- tributed TD(0) with linear function approximation on multi-agent rein- forcement learning,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8e6f3e1-7422-4203-a594-0b1abf1f6707 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning A communication-efficient multi-agent actor-critic algorithm for distributed reinforcement learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f077a0c-e574-4603-a906-f1c42d31a642 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning A multiagent off-policy actor-critic algorithm for distributed reinforcement learning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80511adf-5798-4cfd-8e50-0e2403b88c5e · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Policy gra- dient methods for reinforcement learning with function approximation,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eff02bf0-e670-4cf1-ae71-ba760acbbc3b · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Neural temporal-difference learning converges to global optima,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2f0bb85-2e08-4b71-89de-54d91f7940a2 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Natural actor-critic,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec6a4d39-bfd7-4793-bcd6-f1b2351f9fd3 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Optimistic policy iteration and natural actor-critic: A unify- ing view and a non-optimality result,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 467a7164-fef9-4ed9-874a-e1682bbbfa63 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Improving sample complexity bounds for (natural) actor-critic algorithms,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db101711-24e5-44fb-8a98-339b9e8d1085 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Constrained consensus and optimization in multi-agent networks,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24c32f37-cd00-40fc-8a47-2b49654536d6 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Nesterov,Introductory Lectures on Convex Optimization, Berlin, Germany: Springer, 2018
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a74724f-e082-4d5f-99b7-ff7ba56f5a29 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Convergence rates for localized actor-critic in networked markov potential games,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f465733a-fdb1-400c-89b7-bcc2b5e76e93 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning On the global convergence rates of softmax policy gradient methods,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d00548e1-bb19-448e-9d0a-4fe31563daf8 · outbound
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning Network topology and communication-computation tradeoffs in decentralized optimization,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.