Pith. sign in

REVIEW 3 cited by

Convergence of stochastic gradient descent under a local Lojasiewicz condition for deep neural networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.09221 v2 pith:KWGXOPCC submitted 2023-04-18 cs.LG math.OCstat.ML

classification cs.LGmath.OCstat.ML
keywords localconvergenceconditiondescentgradientnetworksneuralpositive
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

We study the convergence of stochastic gradient descent (SGD) for non-convex objective functions. We establish the local convergence with positive probability under the local \L{}ojasiewicz condition introduced by Chatterjee in \cite{chatterjee2022convergence} and an additional local structural assumption of the loss function landscape. A key component of our proof is to ensure that the whole trajectories of SGD stay inside the local region with a positive probability. We also provide examples of neural networks with finite widths such that our assumptions hold.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Safe Start: Configuring Optimization Algorithms for Decision-Making under Extreme Risks

    math.OC 2026-08 conditional novelty 7.0 of 10

    Safe-start initialization, combined with efficient gradient estimators, guarantees sub-exponential sample complexity for SGD in rare-event optimization; without it, exponential complexity can be unavoidable.

  2. Convergence of Stochastic Gradient Methods for Wide Two-Layer Physics-Informed Neural Networks for the Poisson Equation

    cs.LG 2025-08 conditional novelty 6.0 of 10

    SGD and stochastic gradient flow are proven to drive the empirical PINN loss for the Poisson equation to zero exponentially in expectation, for sufficiently wide two-layer networks.

  3. From Sublinear to Linear: Local Convergence in Finite-Width Networks via Locally Polyak-Lojasiewicz Regions

    stat.ML 2025-07 conditional novelty 4.0 of 10

    Local NTK positivity plus Lipschitz stability gives a local Polyak-Lojasiewicz constant lambda0 minus L_Theta times the region radius, yielding linear gradient descent convergence whenever the iterates stay in the LQC...

Pith tools