DDN, a double distillation network for cooperative MARL, claims to eliminate CTDE inherent error via global-to-local knowledge distillation and an RND-style intrinsic reward, with better SMAC and Predator-Prey win rates.
Leibo, Karl Tuyls, and Thore Grae- pel
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.MA 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Double Distillation Network for Multi-Agent Reinforcement Learning
DDN, a double distillation network for cooperative MARL, claims to eliminate CTDE inherent error via global-to-local knowledge distillation and an RND-style intrinsic reward, with better SMAC and Predator-Prey win rates.