paper-with-me

Papers

Adaptive Consensus Gradients Aggregation for Scaled Distributed Training

2024-11-06 · Yoni Choukroun, Shlomi Azoulay, Pavel Kisilev

Distributed machine learning has recently become a critical paradigm for training large models on vast datasets. We examine the stochastic optimization problem for deep learning within synchronous parallel computing environments under communication constraints. While averaging distributed gradients is the most widely used method for gradient estimation, whether this is the optimal strategy remains an open question. In this work, we analyze the distributed gradient aggregation process through the lens of subspace optimization. By formulating the aggregation problem as an objective-aware subspace optimization problem, we derive an efficient weighting scheme for gradients, guided by subspace coefficients. We further introduce subspace momentum to accelerate convergence while maintaining statistical unbiasedness in the aggregation. Our method demonstrates improved performance over the ubiquitous gradient averaging on multiple MLPerf tasks while remaining extremely efficient in both communicational and computational complexity.

📄 PDF Abstract BibTeX arXiv:2411.03742

Code (1)

yonilc/adacons 공식 구현 pytorch

Tasks

Stochastic Optimization

Similar Papers 제목 키워드 기반

Distributed Average Consensus via Noisy and Non-Coherent Over-the-Air Aggregation

2024-03-11 · Huiwen Yang, Xiaomeng Chen, Lingying Huang, Subhrakanti Dey 외

Over-the-air aggregation has attracted widespread attention for its potential advantages in task-oriented applications, such as distributed sensing, learning, and consensus. In this paper, we develop a communication-effi…

Cooperative guidance of multiple missiles: a hybrid co-evolutionary approach

2022-08-15 · Xuejing Lan, Junda Chen, Zhijia Zhao, Tao Zou

Cooperative guidance of multiple missiles is a challenging task with rigorous constraints of time and space consensus, especially when attacking dynamic targets. In this paper, the cooperative guidance task is described …

continuous-controlContinuous Control

Time-Varying and Nonlinearly Scaled Consensus of Multiagent Systems: A Generic Attracting Law Approach

2020-08-23

This paper presents the design and analysis of the finite/fixed-time scaled consensus for multiagent systems. A study on a generic attracting law, the certain classes of nonlinear systems that admit attractors with finit…

Distributed Consensus Optimization with Consensus ALADIN

2025-03-21 · Xu Du, Jingzhe Wang

TThe paper proposes the Consensus Augmented Lagrange Alternating Direction Inexact Newton (Consensus ALADIN) algorithm, a novel approach for solving distributed consensus optimization problems (DC). Consensus ALADIN allo…

Computational Efficiency

LAPA-based Dynamic Privacy Optimization for Wireless Federated Learning in Heterogeneous Environments

2025-05-26 · Pengcheng Sun, Erwu Liu, Wei Ni, Rui Wang 외

Federated Learning (FL) is a distributed machine learning paradigm based on protecting data privacy of devices, which however, can still be broken by gradient leakage attack via parameter inversion techniques. Differenti…

Federated Learning