paper-with-me

홈 › Papers

Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points

2025-09-21 · Naoya Yamamoto, Juno Kim, Taiji Suzuki arxiv

Wasserstein gradient flow (WGF) is a common method to perform optimization over the space of probability measures. While WGF is guaranteed to converge to a first-order stationary point, for nonconvex functionals the converged solution does not necessarily satisfy the second-order optimality condition; i.e., it could converge to a saddle point. In this work, we propose a new algorithm for probability measure optimization, perturbed Wasserstein gradient flow (PWGF), that achieves second-order optimality for general nonconvex objectives. PWGF enhances WGF by injecting noisy perturbations near saddle points via a Gaussian process-based scheme. By pushing the measure forward along a random vector field generated from a Gaussian process, PWGF helps the solution escape saddle points efficiently by perturbing the solution towards the smallest eigenvalue direction of the Wasserstein Hessian. We theoretically derive the computational complexity for PWGF to achieve a second-order stationary point. Furthermore, we prove that PWGF converges to a global optimum in polynomial time for strictly benign objectives.

📄 PDF Abstract BibTeX arXiv:2509.16974

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Saddle Points Toward Global Minima: A Newton-Type Method on Wasserstein Space

2026-05-18 · Razvan-Andrei Lascu, Taiji Suzuki arxiv

We study the minimization of non-convex functionals over the Wasserstein space. While recent work has showed that perturbed Wasserstein gradient methods can avoid saddle points for benign landscapes, existing approaches …

Multi-Headed Transformer Architectures as Time-dependent Wasserstein Gradient Flows

2026-05-15 · Alex Massucco, Leonardo Del Grande, Marcello Carioni, Christoph Brune 외 arxiv

In recent years, transformer architectures have revolutionized the field of language processing, opening the door to previously unforeseen possibilities. However, from a theoretical point of view, the mathematical models…

Efficiently escaping saddle points on manifolds

2019-06-10 · NeurIPS 2019 12 · Chris Criscitiello, Nicolas Boumal

Smooth, non-convex optimization problems on Riemannian manifolds occur in machine learning as a result of orthonormality, rank or positivity constraints. First- and second-order necessary optimality conditions state that…

Low-Rank Matrix CompletionMatrix Completion

Accelerated Information Gradient flow

2019-09-04 · Yifei Wang, Wuchen Li

We present a framework for Nesterov's accelerated gradient flows in probability space to design efficient mean-field Markov chain Monte Carlo (MCMC) algorithms for Bayesian inverse problems. Here four examples of informa…

Bayesian Inference

Neural Wasserstein Gradient Flows for Maximum Mean Discrepancies with Riesz Kernels

2023-01-27 · Fabian Altekrüger, Johannes Hertrich, Gabriele Steidl

Wasserstein gradient flows of maximum mean discrepancy (MMD) functionals with non-smooth Riesz kernels show a rich structure as singular measures can become absolutely continuous ones and conversely. In this paper we con…