paper-with-me

홈 › Papers

Efficient Federated Fine-Tuning of Large Language Models with Layer Dropout

2025-03-13 · Shilong Wang, Jianchun Liu, Hongli Xu, Jiaming Yan, Xianjun Gao

Fine-tuning plays a crucial role in enabling pre-trained LLMs to evolve from general language comprehension to task-specific expertise. To preserve user data privacy, federated fine-tuning is often employed and has emerged as the de facto paradigm. However, federated fine-tuning is prohibitively inefficient due to the tension between LLM complexity and the resource constraint of end devices, incurring unaffordable fine-tuning overhead. Existing literature primarily utilizes parameter-efficient fine-tuning techniques to mitigate communication costs, yet computational and memory burdens continue to pose significant challenges for developers. This work proposes DropPEFT, an innovative federated PEFT framework that employs a novel stochastic transformer layer dropout method, enabling devices to deactivate a considerable fraction of LLMs layers during training, thereby eliminating the associated computational load and memory footprint. In DropPEFT, a key challenge is the proper configuration of dropout ratios for layers, as overhead and training performance are highly sensitive to this setting. To address this challenge, we adaptively assign optimal dropout-ratio configurations to devices through an exploration-exploitation strategy, achieving efficient and effective fine-tuning. Extensive experiments show that DropPEFT can achieve a 1.3-6.3\times speedup in model convergence and a 40%-67% reduction in memory footprint compared to state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2503.10217

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuning

Methods 이 논문이 사용한 방법론

Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

Sequential Compression Layers for Efficient Federated Learning in Foundational Models

2024-12-09 · Navyansh Mahla, Sunny Gupta, Amit Sethi

Federated Learning (FL) has gained popularity for fine-tuning large language models (LLMs) across multiple nodes, each with its own private data. While LoRA has been widely adopted for parameter efficient federated fine-…

Federated Learningparameter-efficient fine-tuning

FedTLU: Federated Learning with Targeted Layer Updates

2024-12-23 · Jong-Ik Park, Carlee Joe-Wong

Federated learning (FL) addresses privacy concerns in training language models by enabling multiple clients to contribute to the training, without sending their data to others. However, non-IID (identically and independe…

Federated LearningLanguage ModelingLanguage Modelling

SplitFT: An Adaptive Federated Split Learning System For LLMs Fine-Tuning

2026-04-29 · Yimeng Shan, Zhaorui Zhang, Sheng Di, Yu Liu 외 arxiv

Federated Split Learning has been identified as an efficient approach to address the computational resource constraints of clients in classical federated learning, while guaranteeing data privacy for distributed model tr…

Federated Learning

Fisher Information-based Efficient Curriculum Federated Learning with Large Language Models

2024-09-30 · Ji Liu, Jiaxiang Ren, Ruoming Jin, Zijie Zhang 외

As a promising paradigm to collaboratively train models with decentralized data, Federated Learning (FL) can be exploited to fine-tune Large Language Models (LLMs). While LLMs correspond to huge size, the scale of the tr…

Federated Learning

FedVLMBench: Benchmarking Federated Fine-Tuning of Vision-Language Models

2025-06-11 · Weiying Zheng, Ziyue Lin, Pengxin Guo, Yuyin Zhou 외

Vision-Language Models (VLMs) have demonstrated remarkable capabilities in cross-modal understanding and generation by integrating visual and textual information. While instruction tuning and parameter-efficient fine-tun…

BenchmarkingFederated Learningparameter-efficient fine-tuningPrivacy Preserving