paper-with-me

홈 › Papers

Magnitude Matters: Fixing SIGNSGD Through Magnitude-Aware Sparsification in the Presence of Data Heterogeneity

2023-02-19 · Richeng Jin, Xiaofan He, Caijun Zhong, Zhaoyang Zhang, Tony Quek, Huaiyu Dai

Communication overhead has become one of the major bottlenecks in the distributed training of deep neural networks. To alleviate the concern, various gradient compression methods have been proposed, and sign-based algorithms are of surging interest. However, SIGNSGD fails to converge in the presence of data heterogeneity, which is commonly observed in the emerging federated learning (FL) paradigm. Error feedback has been proposed to address the non-convergence issue. Nonetheless, it requires the workers to locally keep track of the compression errors, which renders it not suitable for FL since the workers may not participate in the training throughout the learning process. In this paper, we propose a magnitude-driven sparsification scheme, which addresses the non-convergence issue of SIGNSGD while further improving communication efficiency. Moreover, the local update scheme is further incorporated to improve the learning performance, and the convergence of the proposed method is established. The effectiveness of the proposed scheme is validated through experiments on Fashion-MNIST, CIFAR-10, and CIFAR-100 datasets.

📄 PDF Abstract BibTeX arXiv:2302.09634

Code (0)

등록된 구현이 없습니다.

Tasks

Federated Learning

Similar Papers 제목 키워드 기반

Enhancing SignSGD: Small-Batch Convergence Analysis and a Hybrid Switching Strategy

2026-04-28 · Haoran Chen, Wentao Wang arxiv

SignSGD compresses each stochastic gradient coordinate to a single bit, offering substantial memory and communication savings, but its 1-bit quantization removes magnitude information and is known to leave a generalizati…

SignSVRG: fixing SignSGD via variance reduction

2023-05-22 · Evgenii Chzhen, Sholom Schechtman

We consider the problem of unconstrained minimization of finite sums of functions. We propose a simple, yet, practical way to incorporate variance reduction techniques into SignSGD, guaranteeing convergence that is simil…

Sparse-SignSGD with Majority Vote for Communication-Efficient Distributed Learning

2023-02-15 · Chanho Park, Namyoon Lee

The training efficiency of complex deep learning models can be significantly improved through the use of distributed optimization. However, this process is often hindered by a large amount of communication cost between w…

Deep LearningDistributed OptimizationQuantization

Is magnitude 'generically continuous' for finite metric spaces?

2025-01-15 · Hirokazu Katsumasa, Emily Roff, Masahiko Yoshinaga

Magnitude is a real-valued invariant of metric spaces which, in the finite setting, can be understood as recording the 'effective number of points' in a space as the scale of the metric varies. Motivated by applications …

Topological Data Analysis

Guided AbsoluteGrad: Magnitude of Gradients Matters to Explanation's Localization and Saliency

2024-04-23 · Jun Huang, Yan Liu

This paper proposes a new gradient-based XAI method called Guided AbsoluteGrad for saliency map explanations. We utilize both positive and negative gradient magnitudes and employ gradient variance to distinguish the impo…