paper-with-me

Papers

Bias Dynamics in BabyLMs: Towards a Compute-Efficient Sandbox for Democratising Pre-Training Debiasing

2026-01-14 · Filip Trhlik, Andrew Caines, Paula Buttery arxiv

Pre-trained language models (LMs) have, over the last few years, grown substantially in both societal adoption and training costs. This rapid growth in size has constrained progress in understanding and mitigating their biases. Since re-training LMs is prohibitively expensive, most debiasing work has focused on post-hoc or masking-based strategies, which often fail to address the underlying causes of bias. In this work, we seek to democratise pre-model debiasing research by using low-cost proxy models. Specifically, we investigate BabyLMs, compact BERT-like models trained on small and mutable corpora that can approximate bias acquisition and learning dynamics of larger models. We show that BabyLMs display closely aligned patterns of intrinsic bias formation and performance development compared to standard BERT models, despite their drastically reduced size. Furthermore, correlations between BabyLMs and BERT hold across multiple intra-model and post-model debiasing methods. Leveraging these similarities, we conduct pre-model debiasing experiments with BabyLMs, replicating prior findings and presenting new insights regarding the influence of gender imbalance and toxicity on bias formation. Our results demonstrate that BabyLMs can serve as an effective sandbox for large-scale LMs, reducing pre-training costs from over 500 GPU-hours to under 30 GPU-hours. This provides a way to democratise pre-model debiasing research and enables faster, more accessible exploration of methods for building fairer LMs.

📄 PDF Abstract BibTeX arXiv:2601.09421

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Child-directed speech facilitates production, not comprehension, in BabyLMs

2026-05-31 · Bastian Bunzeck, Sina Zarrieß arxiv

Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on comprehension and not production, which is central to usage-based th…

Language Acquisition

BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context

2025-01-07 · Alexis Matzopoulos, Charl Hendriks, Hishaam Mahomed, Francois Meyer

The BabyLM challenge called on participants to develop sample-efficient language models. Submissions were pretrained on a fixed English corpus, limited to the amount of words children are exposed to in development (<100m…

Language ModellingNERPOSPOS Tagging+1

The NLP Sandbox: an efficient model-to-data system to enable federated and unbiased evaluation of clinical NLP models

2022-06-28 · Yao Yan, Thomas Yu, Kathleen Muenzen, Sijia Liu 외

Objective The evaluation of natural language processing (NLP) models for clinical text de-identification relies on the availability of clinical notes, which is often restricted due to privacy concerns. The NLP Sandbox is…

De-identification

Computer Environments Elicit General Agentic Intelligence in LLMs

2026-01-22 · Daixuan Cheng, Shaohan Huang, Yuxian Gu, Huatong Song 외 arxiv

Agentic intelligence in large language models (LLMs) requires not only model intrinsic capabilities but also interactions with external environments. Equipping LLMs with computers now represents a prevailing trend. Howev…

Long-Context UnderstandingInstruction Following

Anti-Malware Sandbox Games

2022-02-28 · Sujoy Sikdar, Sikai Ruan, Qishen Han, Paween Pitimanaaree 외

We develop a game theoretic model of malware protection using the state-of-the-art sandbox method, to characterize and compute optimal defense strategies for anti-malware. We model the strategic interaction between devel…