Parameter-Free Accelerated Gradient Descent for Nonconvex Minimization

Naoki Marumo,Akiko Takeda
DOI: https://doi.org/10.1137/22m1540934
IF: 2.763
2024-06-19
SIAM Journal on Optimization
Abstract:SIAM Journal on Optimization, Volume 34, Issue 2, Page 2093-2120, June 2024. We propose a new first-order method for minimizing nonconvex functions with a Lipschitz continuous gradient and Hessian. The proposed method is an accelerated gradient descent with two restart mechanisms and finds a solution where the gradient norm is less than [math] in [math] function and gradient evaluations. Unlike existing first-order methods with similar complexity bounds, our algorithm is parameter-free because it requires no prior knowledge of problem-dependent parameters, e.g., the Lipschitz constants and the target accuracy [math]. The main challenge in achieving this advantage is estimating the Lipschitz constant of the Hessian using only first-order information. To this end, we develop a new Hessian-free analysis based on two technical inequalities: a Jensen-type inequality for gradients and an error bound for the trapezoidal rule. Several numerical results illustrate that the proposed method performs comparably to existing algorithms with similar complexity bounds, even without parameter tuning.
mathematics, applied
What problem does this paper attempt to address?