paper-with-me

Papers

A new accelerated gradient method inspired by continuous-time perspective

2021-01-01 · Yasong Feng, Weiguo Gao

Nesterov's accelerated method are widely used in problems with machine learning background including deep learning. To give more insight about the acceleration phenomenon, an ordinary differential equation was obtained from Nesterov's accelerated method by taking step sizes approaching zero, and the relationship between Nesterov's method and the differential equation is still of research interest. In this work, we give the precise order of the iterations of Nesterov's accelerated method converging to the solution of derived differential equation as step sizes go to zero. We then present a new accelerated method with higher order. The new method is more stable than ordinary method for large step size and converges faster. We further apply the new method to matrix completion problem and show its better performance through numerical experiments.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Matrix Completion

Similar Papers 제목 키워드 기반

A More Stable Accelerated Gradient Method Inspired by Continuous-Time Perspective

2021-12-09 · Yasong Feng, Weiguo Gao

Nesterov's accelerated gradient method (NAG) is widely used in problems with machine learning background including deep learning, and is corresponding to a continuous-time differential equation. From this connection, the…

Matrix Completion

A Variational Perspective on Accelerated Methods in Optimization

2016-03-14 · Andre Wibisono, Ashia C. Wilson, Michael. I. Jordan

Accelerated gradient methods play a central role in optimization, achieving optimal rates in many settings. While many generalizations and extensions of Nesterov's original acceleration method have been proposed, it is n…

Accelerated Gradient Descent Escapes Saddle Points Faster than Gradient Descent

2017-11-28 · Chi Jin, Praneeth Netrapalli, Michael. I. Jordan

Nesterov's accelerated gradient descent (AGD), an instance of the general family of "momentum methods", provably achieves faster convergence rate than gradient descent (GD) in the convex setting. However, whether these m…

Generalized Continuous-Time Models for Nesterov's Accelerated Gradient Methods

2024-09-02 · Chanwoong Park, Youngchae Cho, Insoon Yang

Recent research has indicated a substantial rise in interest in understanding Nesterov's accelerated gradient methods via their continuous-time models. However, most existing studies focus on specific classes of Nesterov…

Provably Correct Learning Algorithms in the Presence of Time-Varying Features Using a Variational Perspective

2019-03-12 · Joseph E. Gaudio, Travis E. Gibson, Anuradha M. Annaswamy, Michael A. Bolender

Features in machine learning problems are often time-varying and may be related to outputs in an algebraic or dynamical manner. The dynamic nature of these machine learning problems renders current higher order accelerat…

BIG-bench Machine Learning