paper-with-me

Papers

Step-Ahead Error Feedback for Distributed Training with Compressed Gradient

2020-08-13 · An Xu, Zhouyuan Huo, Heng Huang

Although the distributed machine learning methods can speed up the training of large deep neural networks, the communication cost has become the non-negligible bottleneck to constrain the performance. To address this challenge, the gradient compression based communication-efficient distributed learning methods were designed to reduce the communication cost, and more recently the local error feedback was incorporated to compensate for the corresponding performance loss. However, in this paper, we will show that a new "gradient mismatch" problem is raised by the local error feedback in centralized distributed training and can lead to degraded performance compared with full-precision training. To solve this critical problem, we propose two novel techniques, 1) step ahead and 2) error averaging, with rigorous theoretical analysis. Both our theoretical and empirical results show that our new methods can handle the "gradient mismatch" problem. The experimental results show that we can even train faster with common gradient compression schemes than both the full-precision training and local error feedback regarding the training epochs and without performance loss.

📄 PDF Abstract BibTeX arXiv:2008.05823

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SA-PEF: Step-Ahead Partial Error Feedback for Efficient Federated Learning

2026-01-28 · Dawit Kiros Redie, Reza Arablouei, Stefan Werner arxiv

Biased gradient compression with error feedback (EF) reduces communication in federated learning (FL), but under non-IID data, the residual error can decay slowly, causing gradient mismatch and stalled progress in the ea…

Federated Learning

Error Feedback under $(L_0,L_1)$-Smoothness: Normalization and Momentum

2024-10-22 · Sarit Khirirat, Abdurakhmon Sadiev, Artem Riabinin, Eduard Gorbunov 외

We provide the first proof of convergence for normalized error feedback algorithms across a wide range of machine learning problems. Despite their popularity and efficiency in training deep neural networks, traditional a…

Improved Convergence in Parameter-Agnostic Error Feedback through Momentum

2025-11-18 · Abdurakhmon Sadiev, Yury Demidovich, Igor Sokolov, Grigory Malinovsky 외 arxiv

Communication compression is essential for scalable distributed training of modern machine learning models, but it often degrades convergence due to the noise it introduces. Error Feedback (EF) mechanisms are widely adop…

Error-feedback stochastic modeling strategy for time series forecasting with convolutional neural networks

2020-02-03 · Xinze Zhang, Kun He, Yukun Bao

Despite the superiority of convolutional neural networks demonstrated in time series modeling and forecasting, it has not been fully explored on the design of the neural network architecture and the tuning of the hyper-p…

Time SeriesTime Series AnalysisTime Series Forecasting

AttentionCode: Ultra-Reliable Feedback Codes for Short-Packet Communications

2022-05-30 · Yulin Shao, Emre Ozfatura, Alberto Perotti, Branislav Popovic 외

Ultra-reliable short-packet communication is a major challenge in future wireless networks with critical applications. To achieve ultra-reliable communications beyond 99.999%, this paper envisions a new interaction-based…