Implicit Bias of Linear Equivariant Networks

Andrew Dienes; Bobak T. Kiani; Hannah Lawrence; Kristian Georgiev

arxiv: 2110.06084 · v3 · pith:IW3PRRF5new · submitted 2021-10-12 · 💻 cs.LG · cs.AI

Implicit Bias of Linear Equivariant Networks

Hannah Lawrence , Kristian Georgiev , Andrew Dienes , Bobak T. Kiani This is my paper

classification 💻 cs.LG cs.AI

keywords g-cnnsbiasgroupsimplicitlinearnetworksneuralarchitectures

0 comments

read the original abstract

Group equivariant convolutional neural networks (G-CNNs) are generalizations of convolutional neural networks (CNNs) which excel in a wide range of technical applications by explicitly encoding symmetries, such as rotations and permutations, in their architectures. Although the success of G-CNNs is driven by their \emph{explicit} symmetry bias, a recent line of work has proposed that the \emph{implicit} bias of training algorithms on particular architectures is key to understanding generalization for overparameterized neural nets. In this context, we show that $L$-layer full-width linear G-CNNs trained via gradient descent for binary classification converge to solutions with low-rank Fourier matrix coefficients, regularized by the $2/L$-Schatten matrix norm. Our work strictly generalizes previous analysis on the implicit bias of linear CNNs to linear G-CNNs over all finite groups, including the challenging setting of non-commutative groups (such as permutations), as well as band-limited G-CNNs over infinite groups. We validate our theorems via experiments on a variety of groups, and empirically explore more realistic nonlinear networks, which locally capture similar regularization patterns. Finally, we provide intuitive interpretations of our Fourier space implicit regularization results in real space via uncertainty principles.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

The Implicit Bias of Depth: From Neural Collapse to Softmax Codes
cs.LG 2026-05 unverdicted novelty 7.0

Depth induces an implicit low-rank bias in deep unconstrained feature models trained with unregularized multiclass cross-entropy, promoting softmax codes over neural collapse via more efficient norm propagation.