paper-with-me

홈 › Papers

EMO: Edge Model Overlays to Scale Model Size in Federated Learning

2025-04-01 · Di wu, Weibo He, Wanglei Feng, Zhenyu Wen, Bin Qian, Blesson Varghese

Federated Learning (FL) trains machine learning models on edge devices with distributed data. However, the computational and memory limitations of these devices restrict the training of large models using FL. Split Federated Learning (SFL) addresses this challenge by distributing the model across the device and server, but it introduces a tightly coupled data flow, leading to computational bottlenecks and high communication costs. We propose EMO as a solution to enable the training of large models in FL while mitigating the challenges of SFL. EMO introduces Edge Model Overlay(s) between the device and server, enabling the creation of a larger ensemble model without modifying the FL workflow. The key innovation in EMO is Augmented Federated Learning (AFL), which builds an ensemble model by connecting the original (smaller) FL model with model(s) trained in the overlay(s) to facilitate horizontal or vertical scaling. This is accomplished through three key modules: a hierarchical activation replay cache to decouple AFL from FL, a convergence-aware communication controller to optimize communication overhead, and an ensemble inference module. Evaluations on a real-world prototype show that EMO improves accuracy by up to 17.77% compared to FL, and reduces communication costs by up to 7.17x and decreases training time by up to 6.9x compared to SFL.

📄 PDF Abstract BibTeX arXiv:2504.00726

Code (0)

등록된 구현이 없습니다.

Tasks

Federated Learningmodel

Similar Papers 제목 키워드 기반

PID-Guided Partial Alignment for Multimodal Decentralized Federated Learning

2026-01-15 · Yanhang Shi, Xiaoyu Wang, Houwei Cao, Jian Li 외 arxiv

Multimodal decentralized federated learning (DFL) must support collaboration among agents that hold different modality subsets and often different model components, while operating over peer-to-peer (P2P) overlays withou…

Federated Learning

Privacy-Enhancing Infant Cry Classification with Federated Transformers and Denoising Regularization

2025-12-15 · Geofrey Owino, Bernard Shibwabo arxiv

Infant cry classification can aid early assessment of infant needs. However, deployment of such solutions is limited by privacy concerns around audio data, sensitivity to background noise, and domain shift across recordi…

Federated Learning

Usable Agent Discovery for Decentralized AI Systems

2026-04-25 · Patrizio Dazzi, Emanuele Carlini, Matteo Mordacchini, Saul Urso arxiv

Large-scale agentic systems run on distributed infrastructures where many software agents share physical hosts and are discovered via peer-to-peer mechanisms. Discovery must handle node-level churn from failures and host…

Scaling Law Analysis in Federated Learning: How to Select the Optimal Model Size?

2025-11-15 · Xuanyu Chen, Nan Yang, Shuai Wang, Dong Yuan arxiv

The recent success of large language models (LLMs) has sparked a growing interest in training large-scale models. As the model size continues to scale, concerns are growing about the depletion of high-quality, well-curat…

Federated Learning

Agglomerative Federated Learning: Empowering Larger Model Training via End-Edge-Cloud Collaboration

2023-12-01 · Zhiyuan Wu, Sheng Sun, Yuwei Wang, Min Liu 외

Federated Learning (FL) enables training Artificial Intelligence (AI) models over end devices without compromising their privacy. As computing tasks are increasingly performed by a combination of cloud, edge, and end dev…

Federated Learning