Pith. sign in

REVIEW 1 cited by

Statistical mechanics of transfer learning in fully-connected networks in the proportional limit

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.07168 v1 pith:BLCB6A2Q submitted 2024-07-09 cond-mat.dis-nn cond-mat.stat-mech

classification cond-mat.dis-nncond-mat.stat-mech
keywords learninglimitproportionalfully-connectedgeneralizationnetworkssizetask
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Transfer learning (TL) is a well-established machine learning technique to boost the generalization performance on a specific (target) task using information gained from a related (source) task, and it crucially depends on the ability of a network to learn useful features. Leveraging recent analytical progress in the proportional regime of deep learning theory (i.e. the limit where the size of the training set $P$ and the size of the hidden layers $N$ are taken to infinity keeping their ratio $\alpha = P/N$ finite), in this work we develop a novel single-instance Franz-Parisi formalism that yields an effective theory for TL in fully-connected neural networks. Unlike the (lazy-training) infinite-width limit, where TL is ineffective, we demonstrate that in the proportional limit TL occurs due to a renormalized source-target kernel that quantifies their relatedness and determines whether TL is beneficial for generalization.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Kernel shape renormalization explains output-output correlations in finite Bayesian one-hidden-layer networks

    cond-mat.dis-nn 2024-12 conditional novelty 4.0 of 10

    Output-output correlations in finite Bayesian one-hidden-layer networks follow the kernel shape renormalization order parameter, with readout weight overlap equal to Q*_ab/λ1.

Pith tools