REVIEW 2 cited by
A simple and practical adaptive trust-region method
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
We present an adaptive trust-region method for unconstrained optimization that allows inexact solutions to the trust-region subproblems. Our method is a simple variant of the classical trust-region method of \citet{sorensen1982newton}. The method achieves the best possible convergence bound up to an additive log factor, for finding an $\epsilon$-approximate stationary point, i.e., $O( \Delta_f L^{1/2} \epsilon^{-3/2}) + \tilde{O}(1)$ iterations where $L$ is the Lipschitz constant of the Hessian, $\Delta_f$ is the optimality gap, and $\epsilon$ is the termination tolerance for the gradient norm. This improves over existing trust-region methods whose worst-case bound is at least a factor of $L$ worse. We compare our performance with state-of-the-art trust-region (TRU) and cubic regularization (ARC) methods from the GALAHAD library on the CUTEst benchmark set on problems with more than 100 variables. We use fewer function, gradient, and Hessian evaluations than these methods. For instance, our algorithm's median number of gradient evaluations is $23$ compared to $36$ for TRU and $29$ for ARC. Compared to the conference version of this paper \cite{hamad2022consistently}, our revised method includes several practical enhancements. These modifications dramatically improved performance, including an order of magnitude reduction in the shifted geometric mean of wall-clock times. We also show it suffices for the second derivatives to be locally Lipschitz to guarantee that either the minimum gradient norm converges to zero or the objective value tends towards negative infinity, even when the iterates diverge.
Forward citations
Cited by 2 Pith papers
-
On the Universality of Simple Trust-Region Algorithms
Classical and modified-ratio trust-region methods reach the optimal O(ε^{-1/(1+ν)}) convex and O(ε^{-(2+ν)/(1+ν)}) nonconvex complexity for any Hölder ν∈[0,1] without knowing ν.
-
Accelerating Trust-Region Methods: An Attempt to Balance Global and Local Efficiency
An accelerated trust-region method with a dual-variable local detector claims O~(ε^{-1/3}) global oracle complexity with quadratic local convergence, but the key estimate-sequence inequality is off by a factor of 8.
Discussion (0). Continue with ORCID to comment.