A systematic Atari benchmark finds that Augmented Random Search delivers the best reward per kilowatt hour, while RecurrentPPO and QR-DQN are the least energy-efficient.
llama-models/models/llama3 1/MODEL CARD.md at main - meta-llama/llama-models,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Greener Deep Reinforcement Learning: Analysis of Energy and Carbon Efficiency Across Atari Benchmarks
A systematic Atari benchmark finds that Augmented Random Search delivers the best reward per kilowatt hour, while RecurrentPPO and QR-DQN are the least energy-efficient.