REVIEW 2 cited by
Adaptive First-and Zeroth-order Methods for Weakly Convex Stochastic Optimization Problems
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this paper, we design and analyze a new family of adaptive subgradient methods for solving an important class of weakly convex (possibly nonsmooth) stochastic optimization problems. Adaptive methods that use exponential moving averages of past gradients to update search directions and learning rates have recently attracted a lot of attention for solving optimization problems that arise in machine learning. Nevertheless, their convergence analysis almost exclusively requires smoothness and/or convexity of the objective function. In contrast, we establish non-asymptotic rates of convergence of first and zeroth-order adaptive methods and their proximal variants for a reasonably broad class of nonsmooth \& nonconvex optimization problems. Experimental results indicate how the proposed algorithms empirically outperform stochastic gradient descent and its zeroth-order variant for solving such optimization problems.
Forward citations
Cited by 2 Pith papers
-
A Parameter-Free and Near-Optimal Zeroth-Order Algorithm for Stochastic Convex Optimization
POEM is a parameter-free stochastic zeroth-order method that adapts both step size and smoothing automatically and reaches near-optimal oracle complexity.
-
Refining Adaptive Zeroth-Order Optimization at Ease
R-AdaZO changes the second-moment update of adaptive zeroth-order optimization to use the smoothed first moment instead of the raw gradient estimate, with a new variance-aware convergence analysis and faster empirical...
Discussion (0). Continue with ORCID to comment.