paper-with-me

홈 › Papers

Higher-Order Newton Methods with Polynomial Work per Iteration

2023-11-10 · Amir Ali Ahmadi, Abraar Chaudhry, Jeffrey Zhang

We present generalizations of Newton's method that incorporate derivatives of an arbitrary order $d$ but maintain a polynomial dependence on dimension in their cost per iteration. At each step, our $d^{\text{th}}$-order method uses semidefinite programming to construct and minimize a sum of squares-convex approximation to the $d^{\text{th}}$-order Taylor expansion of the function we wish to minimize. We prove that our $d^{\text{th}}$-order method has local convergence of order $d$. This results in lower oracle complexity compared to the classical Newton method. We show on numerical examples that basins of attraction around local minima can get larger as $d$ increases. Under additional assumptions, we present a modified algorithm, again with polynomial cost per iteration, which is globally convergent and has local convergence of order $d$.

📄 PDF Abstract BibTeX arXiv:2311.06374

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

An Accelerated Newton-GMRES Method for Multilinear PageRank

2025-09-27 · Maryam Boubekraoui, Ridwane Tahiri arxiv

Modeling complex multiway relationships in large-scale networks is becoming more and more challenging in data science. The multilinear PageRank problem, arising naturally in the study of higher-order Markov chains, is a …

Recommendation Systems

Sums of Separable and Quadratic Polynomials

2021-05-11 · Amir Ali Ahmadi, Cemil Dibek, Georgina Hall

We study separable plus quadratic (SPQ) polynomials, i.e., polynomials that are the sum of univariate polynomials in different variables and a quadratic polynomial. Motivated by the fact that nonnegative separable and no…

Multiclass Neural Network Minimization via Tropical Newton Polytope Approximation

2020-01-01 · ICML 2020 1 · Georgios Smyrnis, Petros Maragos

The field of tropical algebra is closely linked with the domain of neural networks with piecewise linear activations, since their output can be described via tropical polynomials in the max-plus semiring. In this work, …

Frugality in second-order optimization: floating-point approximations for Newton's method

2025-11-20 · Giuseppe Carrino, Elena Loli Piccolomini, Elisa Riccietti, Theo Mary arxiv

Minimizing loss functions is central to machine-learning training. Although first-order methods dominate practical applications, higher-order techniques such as Newton's method can deliver greater accuracy and faster con…

Adapting Newton's Method to Neural Networks through a Summary of Higher-Order Derivatives

2023-12-06 · Pierre Wolinski

When training large models, such as neural networks, the full derivatives of order 2 and beyond are usually inaccessible, due to their computational cost. This is why, among the second-order optimization methods, it is v…

Second-order methods