Pith. sign in

REVIEW 1 cited by

Characterizing signal propagation to close the performance gap in unnormalized ResNets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2101.08692 v2 pith:ZPHEP7QS submitted 2021-01-21 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords activationresnetssignaltoolsanalysisbatchnetworksnormalization
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Batch Normalization is a key component in almost all state-of-the-art image classifiers, but it also introduces practical challenges: it breaks the independence between training examples within a batch, can incur compute and memory overhead, and often results in unexpected bugs. Building on recent theoretical analyses of deep ResNets at initialization, we propose a simple set of analysis tools to characterize signal propagation on the forward pass, and leverage these tools to design highly performant ResNets without activation normalization layers. Crucial to our success is an adapted version of the recently proposed Weight Standardization. Our analysis tools show how this technique preserves the signal in networks with ReLU or Swish activation functions by ensuring that the per-channel activation means do not grow with depth. Across a range of FLOP budgets, our networks attain performance competitive with the state-of-the-art EfficientNets on ImageNet.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. NAT: Learning to Attack Neurons for Enhanced Adversarial Transferability

    cs.CV 2025-08 conditional novelty 6.0 of 10

    NAT trains per-neuron adversarial generators that each disrupt one mid-layer neuron, improving cross-model and cross-domain attack transferability over embedding-level baselines.

Pith tools