REVIEW 3 cited by
Normalization Techniques in Training DNNs: Methodology, Analysis and Application
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Normalization Techniques in Training DNNs: Methodology, Analysis and Application
read the original abstract
Normalization techniques are essential for accelerating the training and improving the generalization of deep neural networks (DNNs), and have successfully been used in various applications. This paper reviews and comments on the past, present and future of normalization methods in the context of DNN training. We provide a unified picture of the main motivation behind different approaches from the perspective of optimization, and present a taxonomy for understanding the similarities and differences between them. Specifically, we decompose the pipeline of the most representative normalizing activation methods into three components: the normalization area partitioning, normalization operation and normalization representation recovery. In doing so, we provide insight for designing new normalization technique. Finally, we discuss the current progress in understanding normalization methods, and provide a comprehensive review of the applications of normalization for particular tasks, in which it can effectively solve the key issues.
Forward citations
Cited by 3 Pith papers
-
Elucidating the Design Space of Diffusion-Based Generative Models
Organizing diffusion model design choices yields SOTA FID of 1.79 on CIFAR-10 with only 35 network evaluations per image and similar gains on ImageNet-64.
-
Error-mitigated quantum state tomography using neural networks
Neural networks trained via supervised learning on simulated noisy measurements can mitigate unknown noise in quantum state tomography for pure and mixed states.
-
MoTIF: A Mode-Structured Tensor Framework for Multi-Parametric Approximation, Super-Resolution and Forecasting of Unsteady Systems
MoTIF uses HOSVD to separate multi-parametric unsteady flow data into modal components, applies GPR for parametric and spatial interpolation and RNN for temporal forecasting, achieving under 2% relative RMS error on l...
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.