Pith. sign in

Paper Citation Record · LEDGER

Value-Decomposition Networks For Cooperative Multi-Agent Learning

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 50 inbound Pith citation observations for arXiv:1706.05296.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1706.05296 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 50 of 50 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:35:33.299036Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T19:30:07.924598Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 21f69b57-3810-4938-b009-ea6ca7c1dfe6 · inbound

Growing Action Spaces cites this paper.

Growing Action Spaces Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-25T13:25:52.327460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T13:24:59.682609Z digest=sha256:31379ecea0b7d83ddb297b9cf9aecebf6eed2e47c0b109f017f1b81c04680c2f

Observation a1b3f358-17d5-4180-8e5e-c2984f5859b1 · inbound

Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning cites this paper.

Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T03:25:20.302611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T03:24:33.788346Z digest=sha256:cfcb1d065ccaefdd726725d81a5a1c44dc9a22abaecb55b5aba498436adc00b4

Observation 8ffdbc16-ac46-4810-a962-23a104ba1e77 · inbound

Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning cites this paper.

Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-23T04:17:30.912295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T04:16:54.807723Z digest=sha256:086d1498da0960f340156487e6512aeabc9648854a66b809e461f405fda528e0

Observation f8e7aae4-f630-4d2e-b547-5a61ff47db72 · inbound

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences cites this paper.

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:12:28.489248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T03:09:43.596964Z digest=sha256:c1f799f93f30d7f669f5a6885628d90cfe57d0e73af01fce313809faf0a3ffb0

Observation 12492a1f-4937-442a-90e6-bed2ae472dfe · inbound

Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies cites this paper.

Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-19T01:06:57.448961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T01:05:17.086399Z digest=sha256:5c46134b4b60aedf60e1fa8c45bcc93ee8b140c715b0ca94402d0637d21a46cb

Observation 1e281c07-29d1-42fe-9fe2-7bc6451d9406 · inbound

Incentivizing Safer Actions in Policy Optimization for Constrained Reinforcement Learning cites this paper.

Incentivizing Safer Actions in Policy Optimization for Constrained Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T19:35:33.299036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:35:33.299036Z digest=sha256:77a3898748f4c131d4b868d1c6dff9630e997bed9b8440242c22983b46ad519f

Observation e985d51c-3488-471a-9f9b-141599e91130 · inbound

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem cites this paper.

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T15:16:32.258980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T15:15:26.114479Z digest=sha256:623f2e7c3b65fb77f2040f84358757ed1453cef9daae129cd724ae195140ccda

Observation 1105309e-5526-400a-97db-020f0efcffa4 · inbound

GLo-MAPPO: Multi-Agent Deep Reinforcement Learning for Energy-Efficient UAV-Assisted LoRa Networks cites this paper.

GLo-MAPPO: Multi-Agent Deep Reinforcement Learning for Energy-Efficient UAV-Assisted LoRa Networks Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-18T15:01:32.091405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T14:58:06.782237Z digest=sha256:c8a53444c5de07481586d20d803ac107ada0010c584b840314fb414d13aa305f

Observation 14cda5d8-ec4c-4832-8604-611d7ca059ba · inbound

Agentic Services Computing cites this paper.

Agentic Services Computing Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 162

Resolution
unresolved
no resolver link, observed 2026-08-04T14:41:50.880629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:41:50.880629Z digest=sha256:84f4d26566abe778d2f732360bf02f7c8cad66d73ae538f10a2a4eae963a4fd7

Observation 69c08d2a-b4a7-4e82-a605-ef7a617a1ca2 · inbound

Enhancing Cloud Network Resilience via a Robust LLM-Empowered Multi-Agent Reinforcement Learning Framework cites this paper.

Enhancing Cloud Network Resilience via a Robust LLM-Empowered Multi-Agent Reinforcement Learning Framework Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-21T16:30:22.077120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T16:28:22.036007Z digest=sha256:b3ccf3256b3c287a912b4607eecceedca918bbaa4de5642e741f3115eedad3a7

Observation 15edaa42-740a-4b9e-83f3-884d5f70fafb · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T09:50:49.356797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T09:47:47.051969Z digest=sha256:76ae41100a1c916409d79fb2980bdefff0fee1eef4f41720a5d9a738f82caf3e

Observation 9c42f75c-e6c3-489c-b309-127b024ff62a · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:40.524856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:40.524856Z digest=sha256:7a57da43a6e25391936c8d0fdd1b1e4bdbc317dda7faa628d38d179c5df6cf96

Observation d369e9fd-59b0-469f-8136-c805f087bafb · inbound

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning cites this paper.

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:00:16.952016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T19:58:46.680878Z digest=sha256:49b9dc2964967457e2e0e49423a1bc35be49bef77f58ab563b049a6891a5e100

Observation 5f59326f-30ca-4bb2-90b4-c31b6a39817a · inbound

Reflective Context Learning: Studying the Optimization Primitives of Context Space cites this paper.

Reflective Context Learning: Studying the Optimization Primitives of Context Space Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:33:16.983709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T20:29:13.761153Z digest=sha256:95615c472032cd15eefbf4a57b75f70f74febce9ff6406e1ccf9ef9477d0aab1

Observation 76fcc335-e7a2-4f58-b66f-37371eae5430 · inbound

Do LLM-derived graph priors improve multi-agent coordination? cites this paper.

Do LLM-derived graph priors improve multi-agent coordination? Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.218625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T06:09:08.100187Z digest=sha256:d72fbf037bddb0d5bacaedf2cb247cbe097373c423ada09ca8c78d1dda8a06cf

Observation 6f33302e-bc9b-478d-b62d-bf7db6c64bc7 · inbound

Superminds Test: Actively Evaluating Collective Intelligence of Agent Society via Probing Agents cites this paper.

Superminds Test: Actively Evaluating Collective Intelligence of Agent Society via Probing Agents Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:26:08.797238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T12:00:07.345611Z digest=sha256:a69a67168830e08a40a3fde7aa81d4f45b0bea259d2e0df91974a58bbf269060

Observation ee6ba16b-75e9-4888-9809-f5e949e3783c · inbound

NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search cites this paper.

NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:51:43.028844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T19:08:50.518219Z digest=sha256:04bcf0f43cc8e9ab4dc41e2d6c96e81190b5a19733a113fc46af0374d0c25819

Observation a2f532c2-b896-48dd-8ded-4a1363a28891 · inbound

LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing cites this paper.

LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:26:16.730070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T05:35:12.809203Z digest=sha256:f9025a8c6fd25d7948bed501addab1a674dcb36e824eefcea2b693c324070295

Observation 45acce93-7cfc-44f8-ba95-1d6f5ed31810 · inbound

LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing cites this paper.

LLM-Enhanced Deep Reinforcement Learning for Task Offloading in Collaborative Edge Computing Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:45:08.156777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T23:40:59.430355Z digest=sha256:a1e5d8db58c49e91060eafb320361c3f358374fff94c3a5577f2b35b18040bc4

Observation b852f61f-c39e-424c-bc24-4bef6d2f7f3d · inbound

Coordination Matters: Evaluation of Cooperative Multi-Agent Reinforcement Learning cites this paper.

Coordination Matters: Evaluation of Cooperative Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:01:12.582517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:34:57.125810Z digest=sha256:0263b919c9b1c03b1ee3da04989f52863e4021d17326ac5ac24b91f86e427d1c

Observation e136a762-18ab-40a2-b523-1171c191a2ee · inbound

Randomness is sometimes necessary for coordination cites this paper.

Randomness is sometimes necessary for coordination Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:50:57.851450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T01:01:38.215613Z digest=sha256:ffe35864a1177d34b9d072ae81891d37cdb5913b19f4d6f70c884c3291f8a15a

Observation 5fb11de7-1fd3-45b2-b691-5b5c14058382 · inbound

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning cites this paper.

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:56:27.364955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T01:31:19.576223Z digest=sha256:5f78d2f3ddeacb2274ec592bfd5eb764d5126149536fe3ea8eb94d5892f1cc9b

Observation 452e39d7-2e7a-4bdc-b9a9-317803e94442 · inbound

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning cites this paper.

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:39:10.750452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T22:34:46.440511Z digest=sha256:e020d1f9b5e489643439f642ef3e59a281b7794b06e26b995750fdc21988951c

Observation e237e516-5749-4150-a7c0-298ca305138b · inbound

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning cites this paper.

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:07:22.504840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T06:02:43.375474Z digest=sha256:202a9f3d699512db77248d2a9a8272d5d4e5eb0355f5b2b7f212760819db500e

Observation 64c626c7-ff75-4079-88e2-56cd99525b7f · inbound

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation cites this paper.

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-14T20:12:55.153793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T20:12:07.920183Z digest=sha256:449fe396cdcef5594e977d3e3c4c82f49a51503efcef8a670ed5df6d3383255d

Observation cc8aaf90-7ccb-4f25-a94b-bc3261434234 · inbound

Quantum Advantage in Multi Agent Reinforcement Learning cites this paper.

Quantum Advantage in Multi Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T01:48:28.094133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T01:47:50.831972Z digest=sha256:b195ce1a6ce253fe002cf01409cca54c40a5cd8684f78eb7560c67d78bf12e09

Observation a80837c3-86c7-40b1-ba9f-4c0048adceed · inbound

Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems cites this paper.

Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 205

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T03:08:58.031657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-15T03:07:38.232966Z digest=sha256:a416e9e67ec43ae70b9797b313011b1b8a8fc78babab092c8cdd513db3863b1a

Observation 6ffb02bc-6e2c-48f0-9f60-ab1c697de211 · inbound

Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems cites this paper.

Beyond Individual Intelligence: Surveying Collaboration, Failure Attribution, and Self-Evolution in LLM-based Multi-Agent Systems Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 206

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T16:52:39.995474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T16:51:13.491389Z digest=sha256:880ff588d780e68a3c5336b60aa759a7365922cc0851cac7c6554ea07592769d

Observation 1b3c80dc-bae5-41a1-9406-483087df6cbb · inbound

Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning cites this paper.

Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T12:23:16.743957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T12:23:11.320751Z digest=sha256:bbb320176a3d7702cb5c96896b8e6d80623fcf177e5810e0257ac9e3dcbcc1bf

Observation dcdf6bfc-4a80-44df-8a81-b0f891c8b2de · inbound

Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning cites this paper.

Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T18:55:00.849117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T18:47:09.501908Z digest=sha256:2fef81a55878957bdad83511bc64c3fb1dbe2808a338d8864c43d1deb48f88f6

Observation 71da555e-6cdf-4216-9200-4bdc9f925f19 · inbound

Metric-Gradient Projection for Stable Multi-Agent Policy Learning cites this paper.

Metric-Gradient Projection for Stable Multi-Agent Policy Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:13:46.750682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T22:13:00.107205Z digest=sha256:5fe14c928f4241dcf81d07526b041f3708c39e3f74937ea8303ba7c08a7bcfdb

Observation aeb66533-00e7-4eb6-8f39-3d53493298e3 · inbound

Adaptive Punishment for Cooperation in Mixed-Motive Games cites this paper.

Adaptive Punishment for Cooperation in Mixed-Motive Games Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:24:39.839569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T12:17:53.847812Z digest=sha256:0702238da631988a902acf0484d34d1e280d810568890774a15fd2077e01e203

Observation bfa75ecb-5ed9-4bc8-94ca-7493b35e5ae2 · inbound

Acting on the Unseen: Communication-Free Collaborative Filtering for Decentralized Multi-Robot Task Allocation cites this paper.

Acting on the Unseen: Communication-Free Collaborative Filtering for Decentralized Multi-Robot Task Allocation Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T21:53:58.955453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T21:53:16.504181Z digest=sha256:91d0a013e4267be34e8027d6dd08b0164f434d8afd7cfd324a1ad4f482d4c0ca

Observation 18b1b1a6-be1f-4d92-b702-fc763a5c766d · inbound

Learning Multi-Agent Communication Protocol: Study on Information Entropy Efficiency in MARL cites this paper.

Learning Multi-Agent Communication Protocol: Study on Information Entropy Efficiency in MARL Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:27:22.514032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T20:24:08.842135Z digest=sha256:27c14aa951bae90f7baf5f57c44c6cbe3f51807578e63dfa1e5ebebb545cbf07

Observation e7abb6e6-dd8c-4203-afa7-0d5971f666d4 · inbound

Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning cites this paper.

Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T21:27:24.460649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T19:46:31.662942Z digest=sha256:184d99d98ca1754e2c432079ac581c8efbf17fe9da39686bbec4928e55432b32

Observation 56d01977-ff13-4a44-845a-cb19c32dbf50 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 208

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:23.989549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:a6dcd81ded2c61d6959751b9e68614bd4a57484cdf95eea8241d534409c34368

Observation fa44bb7c-ffa3-49f1-a1b9-97360e57a3a9 · inbound

Continual Quadruped Robots Coordination via Semantic Skill Discovery cites this paper.

Continual Quadruped Robots Coordination via Semantic Skill Discovery Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-02T21:37:25.176584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T19:37:53.168569Z digest=sha256:d8a29965d1ce0ae3a9a0c82a7c0a16a68c58ebb9d52c35a8590e3e00964e264f

Observation 1e4778a3-e938-4fcf-83b5-7259f5ee6745 · inbound

Discovering Interpretable Multi-Parameter Control Policies for Evolutionary Algorithms Using Deep Reinforcement Learning cites this paper.

Discovering Interpretable Multi-Parameter Control Policies for Evolutionary Algorithms Using Deep Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-07-03T00:07:28.189107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T17:29:31.604234Z digest=sha256:8a7008bbdc03f8f1852071c5cf370cc08a4945af3a76d3a889399133eef0234d

Observation 949d17b1-f280-41b2-bd1b-e4958f039b89 · inbound

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria cites this paper.

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T08:47:50.446277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T10:42:17.852996Z digest=sha256:9203edd39ed37c0be57542d23d38d7291cc175778ae085a5e3a6beee6ef19316

Observation ae136284-df7c-49b7-a5ba-f44d31ef2b66 · inbound

Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning cites this paper.

Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-13T07:38:02.368884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:38:02.368884Z digest=sha256:467a3433cb8d72862304e4f60b36046de3ba573478c8289be55f523d979e62ad

Observation 4465ad7a-0bcf-47f7-b262-5bee3fc3d38d · inbound

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse cites this paper.

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:30:07.925797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-25T21:14:31.778575Z digest=sha256:2ced2a4bfaafd95fa7ae96b8e964350ad958dd9cc0cce6c1c51e7b85eafbf2d3

Observation 759da76a-67bf-4e57-b8e4-f9bbac83ad50 · inbound

Revisiting Action Factorization for Complex Action Spaces cites this paper.

Revisiting Action Factorization for Complex Action Spaces Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-04T12:59:52.037528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T05:44:20.417806Z digest=sha256:35c2a2debfb50f1c46e2b8ca8f77e7970e3f08b5e40939a0558d88a22927d254

Observation 859cb2d1-4bcd-4f42-a2a5-1cdf08954008 · inbound

Hierarchical Reinforcement Learning in StarCraft Micromanagement with Influence Maps and Cluster-based Scripts cites this paper.

Hierarchical Reinforcement Learning in StarCraft Micromanagement with Influence Maps and Cluster-based Scripts Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:04:21.512368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T06:00:05.471833Z digest=sha256:866af40703b6899c658f545dd2d86430e205086c8ce4de872f03308f06fbad0f

Observation 48619e02-3662-4db1-b1fd-876bfb5f1057 · inbound

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration cites this paper.

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-06-30T07:24:22.533935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T06:15:29.120280Z digest=sha256:6e8550ae749b8773c846005ce4dd21c987028da55c334fcee3c81637398c9043

Observation 68dc8287-d031-46ba-9ef5-e62d50d80f9c · inbound

Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas cites this paper.

Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T14:47:06.125148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T14:47:06.125148Z digest=sha256:68132680100483a6f55d953530d6500824668cae7118bcf8282fd3d39b285532

Observation fc67145e-d63a-487f-a903-ef1664593bb2 · inbound

Action-Factored Multi-Agent Reinforcement Learning for Scalable Quantum Device Tuning cites this paper.

Action-Factored Multi-Agent Reinforcement Learning for Scalable Quantum Device Tuning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-13T03:08:01.590659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:08:01.590659Z digest=sha256:9214bb5a7cb78f289b022d660cda952381260a9115d65820b98ea92156b67664

Observation ef06af3e-8852-4a78-9188-b2c145bad3b3 · inbound

Value-Aware Prediction for Robust Multi-Agent Coordination Under Communication Loss cites this paper.

Value-Aware Prediction for Robust Multi-Agent Coordination Under Communication Loss Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T16:41:59.158511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:41:59.158511Z digest=sha256:05b0f6a08d92e19db5450569c582cb4c2a8c909fa6724c1dd0f8c32381d12c3a

Observation 7612fb09-f881-420a-8cd6-d4d5effcc29a · inbound

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space cites this paper.

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T15:04:29.757511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:04:29.757511Z digest=sha256:2ce36c2ac477c64df7d4a56426a5b82a5d30a28f7e0c88f506ab216bc6422bb8

Observation 335a733d-2574-476b-8628-e046ddf9749b · inbound

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning cites this paper.

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T01:39:36.573052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:39:36.573052Z digest=sha256:45c1d9242711935fff6f1025382abaa2d71c680f59b1e830d8fc14b73889e9a3

Observation 4868a3a0-0b74-4a6d-be4f-e7448f881c8b · inbound

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation cites this paper.

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T21:55:17.256914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T21:55:17.256914Z digest=sha256:4068df6b3aa8737ac4d428ac32a8d3311697c858e6b76f29720b5190ca9b53cc