paper-with-me

홈 › Papers

Fundamental Limits of Communication Efficiency for Model Aggregation in Distributed Learning: A Rate-Distortion Approach

2022-06-28 · Naifu Zhang, Meixia Tao, Jia Wang, Fan Xu

One of the main focuses in distributed learning is communication efficiency, since model aggregation at each round of training can consist of millions to billions of parameters. Several model compression methods, such as gradient quantization and sparsification, have been proposed to improve the communication efficiency of model aggregation. However, the information-theoretic minimum communication cost for a given distortion of gradient estimators is still unknown. In this paper, we study the fundamental limit of communication cost of model aggregation in distributed learning from a rate-distortion perspective. By formulating the model aggregation as a vector Gaussian CEO problem, we derive the rate region bound and sum-rate-distortion function for the model aggregation problem, which reveals the minimum communication rate at a particular gradient distortion upper bound. We also analyze the communication cost at each iteration and total communication cost based on the sum-rate-distortion function with the gradient statistics of real-world datasets. It is found that the communication gain by exploiting the correlation between worker nodes is significant for SignSGD, and a high distortion of gradient estimator can achieve low total communication cost in gradient compression.

📄 PDF Abstract BibTeX arXiv:2206.13984

Code (0)

등록된 구현이 없습니다.

Tasks

Model CompressionQuantization

Similar Papers 제목 키워드 기반

Information-Theoretic Decentralized Secure Aggregation with Passive Collusion Resilience

2025-08-01 · Xiang Zhang, Zhou Li, Shuangyang Li, Kai Wan 외 arxiv

In decentralized federated learning (FL), multiple clients collaboratively learn a shared machine learning (ML) model by leveraging their privately held datasets distributed across the network, through interactive exchan…

Federated Learning

3PC: Three Point Compressors for Communication-Efficient Distributed Training and a Better Theory for Lazy Aggregation

2022-02-02 · Peter Richtárik, Igor Sokolov, Ilyas Fatkhullin, Elnur Gasanov 외

We propose and study a new class of gradient communication mechanisms for communication-efficient training -- three point compressors (3PC) -- as well as efficient distributed nonconvex optimization algorithms that can t…

Fundamental Limits of Hierarchical Secure Aggregation with Cyclic User Association

2025-03-06 · Xiang Zhang, Zhou Li, Kai Wan, Hua Sun 외

Secure aggregation is motivated by federated learning (FL) where a cloud server aims to compute an averaged model (i.e., weights of deep neural networks) of the locally-trained models of numerous clients, while adhering …

Distributed ComputingFederated Learning

Federated Attention: A Distributed Paradigm for Collaborative LLM Inference over Edge Networks

2025-11-04 · Xiumei Deng, Zehui Xiong, Binbin Chen, Dong In Kim 외 arxiv

Large language models (LLMs) are proliferating rapidly at the edge, delivering intelligent capabilities across diverse application scenarios. However, their practical deployment in collaborative scenarios confronts funda…

Computational Efficiency

The Fundamental Price of Secure Aggregation in Differentially Private Federated Learning

2022-03-07 · Wei-Ning Chen, Christopher A. Choquette-Choo, Peter Kairouz, Ananda Theertha Suresh

We consider the problem of training a $d$ dimensional model with distributed differential privacy (DP) where secure aggregation (SecAgg) is used to ensure that the server only sees the noisy sum of $n$ model updates in e…

Federated Learning