Pith. sign in

REVIEW 2 cited by

Formation of Representations in Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.03006 v2 pith:44XS2DGS submitted 2024-10-03 cs.LG cond-mat.dis-nn

classification cs.LGcond-mat.dis-nn
keywords neuralrepresentationsnetworksalignmentcanonicalemergenceformationhypothesis
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Understanding neural representations will help open the black box of neural networks and advance our scientific understanding of modern AI systems. However, how complex, structured, and transferable representations emerge in modern neural networks has remained a mystery. Building on previous results, we propose the Canonical Representation Hypothesis (CRH), which posits a set of six alignment relations to universally govern the formation of representations in most hidden layers of a neural network. Under the CRH, the latent representations (R), weights (W), and neuron gradients (G) become mutually aligned during training. This alignment implies that neural networks naturally learn compact representations, where neurons and weights are invariant to task-irrelevant transformations. We then show that the breaking of CRH leads to the emergence of reciprocal power-law relations between R, W, and G, which we refer to as the Polynomial Alignment Hypothesis (PAH). We present a minimal-assumption theory proving that the balance between gradient noise and regularization is crucial for the emergence of the canonical representation. The CRH and PAH lead to an exciting possibility of unifying major key deep learning phenomena, including neural collapse and the neural feature ansatz, in a single framework.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Grounding Functional Similarity by Invariance-Aware Model Stitching

    cs.LG 2025-05 conditional novelty 6.0 of 10

    FuLA, a task-agnostic stitching objective that aligns intermediate features through the frozen end network, is claimed to be a more reliable functional similarity metric than task-based stitching.

  2. Ubiquity of Emergent Hebbian Dynamics in Regularized Learning

    cs.LG 2025-05 conditional novelty 6.0 of 10

    L2 weight decay generically makes many learning rules look Hebbian near stationarity, and added noise can make them look anti-Hebbian.

Pith tools