Pith. sign in

REVIEW 1 cited by

Optimization Dynamics of Equivariant and Augmented Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.13458 v5 pith:PWZLSDGS submitted 2023-03-23 cs.LG math.OC

classification cs.LGmath.OC
keywords equivariantlayersaugmenteddataadmissibleanalysisevenmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We investigate the optimization of neural networks on symmetric data, and compare the strategy of constraining the architecture to be equivariant to that of using data augmentation. Our analysis reveals that that the relative geometry of the admissible and the equivariant layers, respectively, plays a key role. Under natural assumptions on the data, network, loss, and group of symmetries, we show that compatibility of the spaces of admissible layers and equivariant layers, in the sense that the corresponding orthogonal projections commute, implies that the sets of equivariant stationary points are identical for the two strategies. If the linear layers of the network also are given a unitary parametrization, the set of equivariant layers is even invariant under the gradient flow for augmented models. Our analysis however also reveals that even in the latter situation, stationary points may be unstable for augmented training although they are stable for the manifestly equivariant models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Data Augmentation and Regularization for Learning Group Equivariance

    stat.ML 2025-02 conditional novelty 5.0 of 10

    With a strong enough non-equivariance penalty, regularized augmented gradient flow converges exponentially fast to the equivariant subspace.

Pith tools