Pith. sign in

Paper Citation Record · LEDGER

An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2409.03052.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.03052 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:06:35.898809Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 782ae683-f35f-4269-8e58-665ee6edf90f · inbound

A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios cites this paper.

A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T21:58:45.695716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:58:45.695716Z digest=sha256:15b80f640b1239ff3f2df617dc6db1556e9c947ad2e7d2f4ecd5ed3f697ed0ff

Observation 75c72700-4ee4-4fc1-90df-51bcb9033149 · inbound

Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning cites this paper.

Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T18:50:49.747276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T18:50:49.747276Z digest=sha256:c67cc88186206539b6707f28f5d360561e09f94253a4536cb09c88b5b77b87ae

Observation 89ca6b7a-aacf-4940-b807-c6b33978a579 · inbound

Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G cites this paper.

Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T17:54:46.259862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:54:46.259862Z digest=sha256:a0064966f8365c89ddd3e131de132fd8ad030a3ccd543e799653c35c94458404

Observation a79c9291-51df-4bc2-8aea-f6f4abf8281d · inbound

Multi-Agent Reinforcement Learning Scheduling to Support Low Latency in Teleoperated Driving cites this paper.

Multi-Agent Reinforcement Learning Scheduling to Support Low Latency in Teleoperated Driving An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:52:40.956366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:52:40.956366Z digest=sha256:61edb6cf5658f9d17ccfbb0f79f755b6b8ef070523ff57f56797852998957a1a

Observation 665d4dda-d0fc-450b-905f-a9f7bf7f5ceb · inbound

Multi-agent Embodied AI: Advances and Future Directions cites this paper.

Multi-agent Embodied AI: Advances and Future Directions An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:16:15.318064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:16:15.318064Z digest=sha256:243ba8878c80e0d958bd0ecf82e60d41f315ae4cddd5d720e87c2cbe3e349e01

Observation 14b5a24a-cb95-4c3c-9211-ae08216c2c98 · inbound

A Multi-Agent Reinforcement Learning Approach for Cooperative Air-Ground-Human Crowdsensing in Emergency Rescue cites this paper.

A Multi-Agent Reinforcement Learning Approach for Cooperative Air-Ground-Human Crowdsensing in Emergency Rescue An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T22:34:11.213822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:34:11.213822Z digest=sha256:78be86f9731fe153ae2097e098af2581327e84c89e49d86750bd3db93174b36b

Observation 1099799c-3f0b-4b0c-9b86-453c0e31561f · inbound

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage cites this paper.

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:12:15.521735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T11:08:51.426575Z digest=sha256:c2aab843adf50eadd94d6564acc4a48c6b6f2cc94869b9c3875ca0294d874c81

Observation a2a49b79-ac74-41a4-9e13-1cc06653ec66 · inbound

ReCoDe: Reinforcement Learning-based Dynamic Constraint Design for Multi-Agent Coordination cites this paper.

ReCoDe: Reinforcement Learning-based Dynamic Constraint Design for Multi-Agent Coordination An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:05:58.333520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:05:58.333520Z digest=sha256:dbb5cbc4f522bb5a78098c8084c2adfe199803778b46ad83b99d6f2740107c3e

Observation d8fa179c-fde7-4f44-b1b1-ab79d3212837 · inbound

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning cites this paper.

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:40:38.686052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:40:38.686052Z digest=sha256:69814b7135f61f1f5a0122279a91ebaa49cd66694f9f5f6fe842f6e1992da510

Observation 44742ba5-51e9-4b7f-932b-1efb0f5ff205 · inbound

Real-time adaptive quantum error correction by model-free multi-agent learning cites this paper.

Real-time adaptive quantum error correction by model-free multi-agent learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-05T10:37:10.405790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:37:10.405790Z digest=sha256:0feb45f8a9987b22dbb720a40439d9debd0031e52d2281ad1099499332e7feb8

Observation dd073fbe-2a90-42f6-972e-65120b48f8ff · inbound

Coupling Smoothed Particle Hydrodynamics with Multi-Agent Deep Reinforcement Learning for Cooperative Control of Point Absorbers cites this paper.

Coupling Smoothed Particle Hydrodynamics with Multi-Agent Deep Reinforcement Learning for Cooperative Control of Point Absorbers An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T11:28:05.024515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:28:05.024515Z digest=sha256:821b0c836dc7840d756d495cb8abb6ae1ce14ae93180cd51d79ade51bfcda8b1

Observation eb300f35-c264-4348-8be0-7189985268ae · inbound

AGMARL-DKS: An Adaptive Graph-Enhanced Multi-Agent Reinforcement Learning for Dynamic Kubernetes Scheduling cites this paper.

AGMARL-DKS: An Adaptive Graph-Enhanced Multi-Agent Reinforcement Learning for Dynamic Kubernetes Scheduling An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:59:59.397066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T11:57:24.591538Z digest=sha256:2837cd155a4c6a9ec219ffcf6f32c6645fcccbe37a543109eff39983e80eb9e0

Observation cb1ac08b-b1da-4f0b-a87a-c074f88c0bf8 · inbound

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks cites this paper.

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:55:58.907161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:55:19.978358Z digest=sha256:7c84b1b00a001854e0de390f15e197ea6535f23c012bea5741565277242f59ed

Observation f127e8cc-0c51-4bce-8faf-5cbd518ea800 · inbound

Do LLM-derived graph priors improve multi-agent coordination? cites this paper.

Do LLM-derived graph priors improve multi-agent coordination? An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.221104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T06:09:08.100187Z digest=sha256:5e61be685ccf1c0591e9f62ebf8982539c967e1424ef68260fb6453963da14a9

Observation 48714f95-240c-41bd-9963-1c7687bce540 · inbound

Cross-Modal Navigation with Multi-Agent Reinforcement Learning cites this paper.

Cross-Modal Navigation with Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:12.757224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T08:43:54.882467Z digest=sha256:578f3a8fe01e65cebefb19e73ab7d9604e76e6b5da959da0b8b1f81cd08d80d4

Observation 630fcd5f-cbcd-4d5f-92f8-603fd71df216 · inbound

ERPPO: Entropy Regularization-based Proximal Policy Optimization cites this paper.

ERPPO: Entropy Regularization-based Proximal Policy Optimization An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:32:51.887848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-14T19:32:24.940497Z digest=sha256:7e3944bdd9e25f2680bfa3ff5600a1b3d1876dcc79adf47076841b1e87539fb9

Observation e0b1f735-c955-477c-9042-6794e1f3754a · inbound

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning cites this paper.

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T20:55:04.121723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T20:52:06.448213Z digest=sha256:864d3e19f1f2e4b97d9fa67ff8372545ffef52666e8b27fea59ee98cd092ff4e

Observation 4bb3849e-ca0e-42b2-9395-bb96262f48b2 · inbound

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming cites this paper.

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T15:27:38.614899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T15:27:22.127896Z digest=sha256:275d92290157d6551547854ad8cd469b0429147ddc4e630fea00c1fe6754b1cb

Observation 7f4be0b3-7adb-41e8-ab15-2a685a9a60aa · inbound

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks cites this paper.

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:43:23.574722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T11:41:30.827427Z digest=sha256:02b21ac4a25ee737e9406152e89ece150985aea48f8806395a312c22fa6aac1e

Observation e61e9a24-c40b-4c47-b376-7d1b108638ee · inbound

CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy cites this paper.

CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-27T10:00:49.182519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T09:55:03.745087Z digest=sha256:7f138a9275dd57145d852eabc56af3f01555ed87bf16c90a0ae6561757d6658b

Observation 5c75ce09-454a-4ecf-8e76-0fb44705bfdc · inbound

Self-CTRL: Self-Consistency Training with Reinforcement Learning cites this paper.

Self-CTRL: Self-Consistency Training with Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:56.006777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T01:38:48.296421Z digest=sha256:6dc5418b7092bdba135f2ed0a5642596f5edf9ab0e6e1319dfd1d5edf507f669

Observation 2fa175ba-26b1-47bc-8294-906142bf628d · inbound

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization cites this paper.

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.037734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-08T01:42:38.113751Z digest=sha256:3a87a7f9a3da1b7f1d80df8f4c491392adef29c6171d19e5c265ca206171cdd3

Observation b6c0562f-ae7e-4f06-b25f-d52cc58b1241 · inbound

Social-spatial dependencies for learning visual navigation cites this paper.

Social-spatial dependencies for learning visual navigation An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-09T10:26:10.921508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-07-09T10:23:00.725686Z digest=sha256:5acdeabba85e0db445d063f705ac14e4773930727f7ede580be572a643aa1709

Observation 02cff74d-13ab-47c2-a7df-cd902bf4b2d1 · inbound

Reinforcement Learning for Delivery Drone-Based Participatory Sensing in Dynamic Environments cites this paper.

Reinforcement Learning for Delivery Drone-Based Participatory Sensing in Dynamic Environments An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T14:08:04.235006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:08:04.235006Z digest=sha256:7f877d19cbeaaf4c065b62c5059e2f8b44371bf1eed49a699c2f8eb7aa13d9b0

Observation c811c081-173a-4935-8d7f-0fccb6dc2993 · inbound

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs cites this paper.

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:52:00.844770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:52:00.844770Z digest=sha256:03c58babfa88894bbf7432956aef2feddb11acb3bd1cbbe227924e3a6638e9c1

Observation 252e686e-8d5a-4464-be9a-213793c6c3fa · inbound

Latent Semantic State Estimation for Reliable Swarming of UAVs under Intermittent Connectivity cites this paper.

Latent Semantic State Estimation for Reliable Swarming of UAVs under Intermittent Connectivity An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:25:48.181291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:25:48.181291Z digest=sha256:6c646c6bca79da1e5369baf87b44dc96157749ebebe9a349ea302e846fca5a8f

Observation 5074cb71-2bf8-4341-9c92-132d71c9fe2e · inbound

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition cites this paper.

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T11:20:20.055396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:20:20.055396Z digest=sha256:3fd54db570ec56aeef012eb781514abd5e9dcdb434c3647b7625483d6186c4d5

Observation afa5d363-c047-481b-9121-1601d3146f91 · inbound

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry cites this paper.

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-16T00:06:35.898809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:06:35.898809Z digest=sha256:cd90e51d0bab86c064cd309704ac4528b8c0d039615b0559ad04dba61b96ed2f