paper-with-me

홈 › Papers

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

2025-09-19 · Baichuan Huang, Ananth Balashankar, Amir Aminifar arxiv

Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competitive performance, particularly in low-data regimes. However, the link between fine-tuning different bias terms (i.e., $\boldsymbol{b}_q$, $\boldsymbol{b}_k$, and $\boldsymbol{b}_v$ in the query, key, or value projections) and downstream performance remains largely unclear to date. In this paper, we investigate the link between fine-tuning $\boldsymbol{b}_q$, $\boldsymbol{b}_k$, and $\boldsymbol{b}_v$ with the performance of the downstream task. Our key finding is that directly fine-tuning $\boldsymbol{b}_v$ generally leads to higher downstream performance in low-data regimes, in comparison to $\boldsymbol{b}_q$ and $\boldsymbol{b}_k$. We extensively evaluate this unique property across a wide range of LLMs spanning encoder-only and decoder-only architectures up to 6.7B parameters (including bias-free LLMs). Our results provide strong evidence for the effectiveness of directly fine-tuning $\boldsymbol{b}_v$ across various downstream tasks. The implementation code is available at https://github.com/whubaichuan/BEFT.

📄 PDF Abstract BibTeX arXiv:2509.15974

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

BERAG: Bayesian Ensemble Retrieval-Augmented Generation for Knowledge-based Visual Question Answering

2026-04-24 · Jinghong Chen, Jingbiao Mei, Guangyu Yang, Bill Byrne arxiv

A common approach to question answering with retrieval-augmented generation (RAG) is to concatenate documents into a single context and pass it to a language model to generate an answer. While simple, this strategy can o…

Visual Question Answering

RobustDebias: Debiasing Language Models using Distributionally Robust Optimization

2026-01-30 · Deep Gandhi, Katyani Singh, Nidhi Hegde arxiv

Pretrained language models have been shown to exhibit biases and social stereotypes. Prior work on debiasing these models has largely focused on modifying embedding spaces during pretraining, which is not scalable for la…

Gender-tuning: Empowering Fine-tuning for Debiasing Pre-trained Language Models

2023-07-20 · Somayeh Ghanbarzadeh, Yan Huang, Hamid Palangi, Radames Cruz Moreno 외

Recent studies have revealed that the widely-used Pre-trained Language Models (PLMs) propagate societal biases from the large unmoderated pre-training corpora. Existing solutions require debiasing training processes and …

Language ModelingLanguage ModellingMasked Language Modeling

BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models

2021-10-16 · ACL ARR October 2021 10 · Anonymous

We show that with small-to-medium training data, fine-tuning only the bias terms (or a subset of the bias terms) of pre-trained BERT models is competitive with (and sometimes better than) fine-tuning the entire model. Fo…

Language ModelingLanguage Modellingparameter-efficient fine-tuning

On Transferability of Bias Mitigation Effects in Language Model Fine-Tuning

2020-10-24 · NAACL 2021 4 · Xisen Jin, Francesco Barbieri, Brendan Kennedy, Aida Mostafazadeh Davani 외

Fine-tuned language models have been shown to exhibit biases against protected groups in a host of modeling tasks such as text classification and coreference resolution. Previous works focus on detecting these biases, re…

coreference-resolutionCoreference ResolutionFairnessHate Speech Detection+6