The paper formalizes hidden Byzantine action overwrites in cooperative multi-agent RL, proves an exact rectangular robust-MDP reduction, shows linear security regret is unavoidable under uninformative attacks, and gives a stage-tied E2D learner with ~O(H^2 S sqrt(AK)) + E[D_K] regret.
A comprehensive survey on multi-agent reinforcement learning for connected and automated vehicles.Sensors, 23(10):4710, 2023
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.LG 1years
2026 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Online Security Learning in Cooperative Multi-Agent Systems under Hidden Byzantine Attacks
The paper formalizes hidden Byzantine action overwrites in cooperative multi-agent RL, proves an exact rectangular robust-MDP reduction, shows linear security regret is unavoidable under uninformative attacks, and gives a stage-tied E2D learner with ~O(H^2 S sqrt(AK)) + E[D_K] regret.