REDA learns per-agent Q-values and uses them as benefit inputs to an optimal assignment mechanism, outperforming IQL, IPPO, COMA, and HAAL on sequential satellite assignment.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.MA 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Multi Agent Reinforcement Learning for Sequential Satellite Assignment Problems
REDA learns per-agent Q-values and uses them as benefit inputs to an optimal assignment mechanism, outperforming IQL, IPPO, COMA, and HAAL on sequential satellite assignment.