paper-with-me

홈 › Papers

Decentralized Riemannian natural gradient methods with Kronecker-product approximations

2023-03-16 · Jiang Hu, Kangkang Deng, Na Li, Quanzheng Li

With a computationally efficient approximation of the second-order information, natural gradient methods have been successful in solving large-scale structured optimization problems. We study the natural gradient methods for the large-scale decentralized optimization problems on Riemannian manifolds, where the local objective function defined by the local dataset is of a log-probability type. By utilizing the structure of the Riemannian Fisher information matrix (RFIM), we present an efficient decentralized Riemannian natural gradient descent (DRNGD) method. To overcome the communication issue of the high-dimension RFIM, we consider a class of structured problems for which the RFIM can be approximated by a Kronecker product of two low-dimension matrices. By performing the communications over the Kronecker factors, a high-quality approximation of the RFIM can be obtained in a low cost. We prove that DRNGD converges to a stationary point with the best-known rate of $\mathcal{O}(1/K)$. Numerical experiments demonstrate the efficiency of our proposed method compared with the state-of-the-art ones. To the best of our knowledge, this is the first Riemannian second-order method for solving decentralized manifold optimization problems.

📄 PDF Abstract BibTeX arXiv:2303.09611

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Natural Gradient Descent 설명 없음

Similar Papers 제목 키워드 기반

A Coordinate-Free Construction of Scalable Natural Gradient

2018-08-30 · Kevin Luk, Roger Grosse

Most neural networks are trained using first-order optimization methods, which are sensitive to the parameterization of the model. Natural gradient descent is invariant to smooth reparameterizations because it is defined…

Decentralized Riemannian Conjugate Gradient Method on the Stiefel Manifold

2023-08-21 · Jun Chen, Haishan Ye, Mengmeng Wang, Tianxin Huang 외

The conjugate gradient method is a crucial first-order optimization method that generally converges faster than the steepest descent method, and its computational cost is much lower than that of second-order methods. How…

Second-order methods

Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds

2025-09-09 · Emre Sahinoglu, Shahin Shahrampour arxiv

We study decentralized online Riemannian optimization over manifolds with possibly positive curvature, going beyond the Hadamard manifold setting. Decentralized optimization techniques rely on a consensus step that is we…

Decentralized Riemannian Gradient Descent on the Stiefel Manifold

2021-02-14 · Shixiang Chen, Alfredo Garcia, Mingyi Hong, Shahin Shahrampour

We consider a distributed non-convex optimization where a network of agents aims at minimizing a global function over the Stiefel manifold. The global function is represented as a finite sum of smooth local functions, wh…

Distributed Optimization

Decentralized Optimization on Compact Submanifolds by Quantized Riemannian Gradient Tracking

2025-06-09 · Jun Chen, Lina Liu, Tianyi Zhu, Yong liu 외

This paper considers the problem of decentralized optimization on compact submanifolds, where a finite sum of smooth (possibly non-convex) local functions is minimized by $n$ agents forming an undirected and connected gr…

Distributed OptimizationQuantization