Pith. sign in

Optimization methods for large-scale machine learning

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.LG 1

years

2019 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

LCA: Loss Change Allocation for Neural Network Training

cs.LG · 2019-09-03 · conditional · novelty 6.0

A per-parameter decomposition of training loss change shows that learning is noisy, with only about half of parameters helping per step, some layers hurting overall, and learning spikes synchronized across layers.

citing papers explorer

Showing 1 of 1 citing paper.

  • LCA: Loss Change Allocation for Neural Network Training cs.LG · 2019-09-03 · conditional · none · ref 3

    A per-parameter decomposition of training loss change shows that learning is noisy, with only about half of parameters helping per step, some layers hurting overall, and learning spikes synchronized across layers.