Pith. sign in

REVIEW

Strong convexity-guided hyper-parameter optimization for flatter losses

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.05025 v1 pith:S2FSNUT6 submitted 2024-02-07 cs.LG

classification cs.LG
keywords strongconvexityhyper-parameterfindflatnesslossoptimizationrelationship
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We propose a novel white-box approach to hyper-parameter optimization. Motivated by recent work establishing a relationship between flat minima and generalization, we first establish a relationship between the strong convexity of the loss and its flatness. Based on this, we seek to find hyper-parameter configurations that improve flatness by minimizing the strong convexity of the loss. By using the structure of the underlying neural network, we derive closed-form equations to approximate the strong convexity parameter, and attempt to find hyper-parameters that minimize it in a randomized fashion. Through experiments on 14 classification datasets, we show that our method achieves strong performance at a fraction of the runtime.

Discussion (0). Continue with ORCID to comment.

Pith tools