paper-with-me

홈 › Papers

Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

2025-04-10 · Alex Warstadt, Aaron Mueller, Leshem Choshen, Ethan Wilcox, Chengxu Zhuang, Juan Ciro, Rafael Mosquera, Bhargavi Paranjape, Adina Williams, Tal Linzen, Ryan Cotterell

Children can acquire language from less than 100 million words of input. Large language models are far less data-efficient: they typically require 3 or 4 orders of magnitude more data and still do not perform as well as humans on many evaluations. These intensive resource demands limit the ability of researchers to train new models and use existing models as developmentally plausible cognitive models. The BabyLM Challenge is a communal effort in which participants compete to optimize language model training on a fixed data budget. Submissions are compared on various evaluation tasks targeting grammatical ability, downstream task performance, and generalization. Participants can submit to up to three tracks with progressively looser data restrictions. From over 30 submissions, we extract concrete recommendations on how best to train data-efficient language models, and on where future efforts should (and perhaps should not) focus. The winning submissions using the LTG-BERT architecture (Samuel et al., 2023) outperformed models trained on trillions of words. Other submissions achieved strong results through training on shorter input sequences or training a student model on a pretrained teacher. Curriculum learning attempts, which accounted for a large number of submissions, were largely unsuccessful, though some showed modest improvements.

📄 PDF Abstract BibTeX arXiv:2504.08165

Code (1)

babylm/evaluation-pipeline 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus

2023-01-27 · Alex Warstadt, Leshem Choshen, Aaron Mueller, Adina Williams 외

We present the call for papers for the BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus. This shared task is intended for participants with an interest in small scale language modeling…

Language AcquisitionLanguage ModelingLanguage ModellingNatural Language Understanding

[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus

2024-04-09 · Leshem Choshen, Ryan Cotterell, Michael Y. Hu, Tal Linzen 외

After last year's successful BabyLM Challenge, the competition will be hosted again in 2024/2025. The overarching goals of the challenge remain the same; however, some of the competition rules will be different. The big …

Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

2024-12-06 · Michael Y. Hu, Aaron Mueller, Candace Ross, Adina Williams 외

The BabyLM Challenge is a community effort to close the data-efficiency gap between human and computational language learners. Participants compete to optimize language model training on a fixed language data budget of 1…

Language ModelingLanguage ModellingQuestion AnsweringVisual Question Answering

Can training neural language models on a curriculum with developmentally plausible data improve alignment with human reading behavior?

2023-11-30 · Aryaman Chobey, Oliver Smith, Anzi Wang, Grusha Prasad

The use of neural language models to model human behavior has met with mixed success. While some work has found that the surprisal estimates from these models can be used to predict a wide range of human neural and behav…

Sentence

Model Merging to Maintain Language-Only Performance in Developmentally Plausible Multimodal Models

2025-10-02 · Ece Takmaz, Lisa Bylinina, Jakub Dotlacil arxiv

State-of-the-art vision-and-language models consist of many parameters and learn from enormous datasets, surpassing the amounts of linguistic data that children are exposed to as they acquire a language. This paper prese…