paper-with-me

홈 › Papers

Deploying Large AI Models on Resource-Limited Devices with Split Federated Learning

2025-04-12 · Xianke Qiang, Hongda Liu, Xinran Zhang, Zheng Chang, Ying-Chang Liang

Large Artificial Intelligence Models (LAMs) powered by massive datasets, extensive parameter scales, and extensive computational resources, leading to significant transformations across various industries. Yet, their practical deployment on resource-limited mobile edge devices is hindered by critical challenges such as data privacy, constrained resources, and high overhead costs. Addressing this gap, this paper proposes a novel framework, named Quantized Split Federated Fine-Tuning Large AI Model (SFLAM). By partitioning the training load between edge devices and servers using a split learning paradigm, SFLAM can facilitate the operation of large models on devices and significantly lowers the memory requirements on edge devices. Additionally, SFLAM incorporates quantization management, power control, and bandwidth allocation strategies to enhance training efficiency while concurrently reducing energy consumption and communication latency. A theoretical analysis exploring the latency-energy trade-off is presented, and the framework's efficacy is validated via comprehensive simulations. The findings indicate that SFLAM achieves superior performance in terms of learning efficiency and scalability compared to conventional methods, thereby providing a valuable approach for enabling advanced AI services in resource-constrained scenarios.

📄 PDF Abstract BibTeX arXiv:2504.09114

Code (0)

등록된 구현이 없습니다.

Tasks

Federated LearningQuantization

Similar Papers 제목 키워드 기반

Split Knowledge Distillation for Large Models in IoT: Architecture, Challenges, and Solutions

2024-12-17 · Zuguang Li, Wen Wu, Shaohua Wu, Qiaohua Lin 외

Large models (LMs) have immense potential in Internet of Things (IoT) systems, enabling applications such as intelligent voice assistants, predictive maintenance, and healthcare monitoring. However, training LMs on edge …

Knowledge DistillationManagement

Dynamic Split Computing for Efficient Deep Edge Intelligence

2022-05-23 · Arian Bakhtiarnia, Nemanja Milošević, Qi Zhang, Dragana Bajović 외

Deploying deep neural networks (DNNs) on IoT and mobile devices is a challenging task due to their limited computational resources. Thus, demanding tasks are often entirely offloaded to edge servers which can accelerate …

Edge-computingHyperparameter Optimization

AutoDiCE: Fully Automated Distributed CNN Inference at the Edge

2022-07-20 · Xiaotian Guo, Andy D. Pimentel, Todor Stefanov

Deep Learning approaches based on Convolutional Neural Networks (CNNs) are extensively utilized and very successful in a wide range of application areas, including image classification and speech recognition. For the exe…

Code Generationimage-classificationImage Classificationspeech-recognition+1

A QoE-Aware Split Inference Accelerating Algorithm for NOMA-based Edge Intelligence

2024-09-25 · Xin Yuan, Ning li, Quan Chen, Wenchao Xu 외

Even the AI has been widely used and significantly changed our life, deploying the large AI models on resource limited edge devices directly is not appropriate. Thus, the model split inference is proposed to improve the …

Efficient Split Learning LSTM Models for FPGA-based Edge IoT Devices

2025-02-12 · Romina Soledad Molina, Vukan Ninkovic, Dejan Vukobratovic, Maria Liz Crespo 외

Split Learning (SL) recently emerged as an efficient paradigm for distributed Machine Learning (ML) suitable for the Internet Of Things (IoT)-Cloud systems. However, deploying SL on resource-constrained edge IoT platform…