paper-with-me

홈 › Papers

Revisiting Optimal Convergence Rate for Smooth and Non-convex Stochastic Decentralized Optimization

2022-10-14 · Kun Yuan, Xinmeng Huang, Yiming Chen, Xiaohan Zhang, Yingya Zhang, Pan Pan

Decentralized optimization is effective to save communication in large-scale machine learning. Although numerous algorithms have been proposed with theoretical guarantees and empirical successes, the performance limits in decentralized optimization, especially the influence of network topology and its associated weight matrix on the optimal convergence rate, have not been fully understood. While (Lu and Sa, 2021) have recently provided an optimal rate for non-convex stochastic decentralized optimization with weight matrices defined over linear graphs, the optimal rate with general weight matrices remains unclear. This paper revisits non-convex stochastic decentralized optimization and establishes an optimal convergence rate with general weight matrices. In addition, we also establish the optimal rate when non-convex loss functions further satisfy the Polyak-Lojasiewicz (PL) condition. Following existing lines of analysis in literature cannot achieve these results. Instead, we leverage the Ring-Lattice graph to admit general weight matrices while maintaining the optimal relation between the graph diameter and weight matrix connectivity. Lastly, we develop a new decentralized algorithm to nearly attain the above two optimal rates under additional mild conditions.

📄 PDF Abstract BibTeX arXiv:2210.07863

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Revisiting Projection-Free Optimization for Strongly Convex Constraint Sets

2018-11-14 · Jarrid Rector-Brooks, Jun-Kun Wang, Barzan Mozafari

We revisit the Frank-Wolfe (FW) optimization under strongly convex constraint sets. We provide a faster convergence rate for FW without line search, showing that a previously overlooked variant of FW is indeed faster tha…

Revisiting Convergence: Shuffling Complexity Beyond Lipschitz Smoothness

2025-07-11 · Qi He, Peiran Yu, Ziyi Chen, Heng Huang arxiv

Shuffling-type gradient methods are favored in practice for their simplicity and rapid empirical performance. Despite extensive development of convergence guarantees under various assumptions in recent years, most requir…

Revisiting Subgradient Method: Complexity and Convergence Beyond Lipschitz Continuity

2023-05-23 · Xiao Li, Lei Zhao, Daoli Zhu, Anthony Man-Cho So

The subgradient method is one of the most fundamental algorithmic schemes for nonsmooth optimization. The existing complexity and convergence results for this method are mainly derived for Lipschitz continuous objective …

Near-Optimal Non-Convex Stochastic Optimization under Generalized Smoothness

2023-02-13 · Zijian Liu, Srikanth Jagabathula, Zhengyuan Zhou

The generalized smooth condition, $(L_{0},L_{1})$-smoothness, has triggered people's interest since it is more realistic in many optimization problems shown by both empirical and theoretical evidence. Two recent works es…

Stochastic Optimization

Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods

2023-12-13 · Zijian Liu, Zhengyuan Zhou

In the past several years, the last-iterate convergence of the Stochastic Gradient Descent (SGD) algorithm has triggered people's interest due to its good performance in practice but lack of theoretical understanding. Fo…