A maximum-entropy mean-field deep Q-network is proposed for joint trajectory, user association, and power control in dense UAV networks.
Large popula tion stochastic dynamic games: Closed-loop McKean-Vlasov syst ems and the Nash certainty equivalence principle,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.SY 1years
2024 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Dynamic Trajectory and Power Control in Ultra-Dense UAV Networks: A Mean-Field Reinforcement Learning Approach
A maximum-entropy mean-field deep Q-network is proposed for joint trajectory, user association, and power control in dense UAV networks.