Pith. sign in

On the sdes and scaling rules for adaptive gradient algorithms,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Energy Consumption in Parallel Neural Network Training

cs.LG · 2025-08-11 · conditional · novelty 6.0

Energy use in data-parallel neural network training grows roughly linearly with GPU hours, but the energy cost per GPU hour varies by model, hardware, and the number of samples and gradient updates per GPU hour.

citing papers explorer

Showing 1 of 1 citing paper.

  • Energy Consumption in Parallel Neural Network Training cs.LG · 2025-08-11 · conditional · none · ref 7

    Energy use in data-parallel neural network training grows roughly linearly with GPU hours, but the energy cost per GPU hour varies by model, hardware, and the number of samples and gradient updates per GPU hour.