WA3C extends A3C with a priority-weighted softmax and a five-term reward, and the paper reports simulated gains in latency, energy, and fairness over six baselines.
A2c-drl: Dynamic scheduling for stochastic edge-cloud environments using a2c and deep reinforcement learning,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.DC 1years
2025 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Adaptive, Efficient and Fair Resource Allocation in Cloud Datacenters leveraging Weighted A3C Deep Reinforcement Learning
WA3C extends A3C with a priority-weighted softmax and a five-term reward, and the paper reports simulated gains in latency, energy, and fairness over six baselines.