MARL-CPC lets decentralized agents learn to send informative messages through a self-supervised reconstruction objective, and outperforms message-as-action baselines in non-cooperative multi-agent tasks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.MA 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Reward-Independent Messaging for Decentralized Multi-Agent Reinforcement Learning
MARL-CPC lets decentralized agents learn to send informative messages through a self-supervised reconstruction objective, and outperforms message-as-action baselines in non-cooperative multi-agent tasks.