paper-with-me

홈 › Papers

Can training neural language models on a curriculum with developmentally plausible data improve alignment with human reading behavior?

2023-11-30 · Aryaman Chobey, Oliver Smith, Anzi Wang, Grusha Prasad

The use of neural language models to model human behavior has met with mixed success. While some work has found that the surprisal estimates from these models can be used to predict a wide range of human neural and behavioral responses, other work studying more complex syntactic phenomena has found that these surprisal estimates generate incorrect behavioral predictions. This paper explores the extent to which the misalignment between empirical and model-predicted behavior can be minimized by training models on more developmentally plausible data, such as in the BabyLM Challenge. We trained teacher language models on the BabyLM "strict-small" dataset and used sentence level surprisal estimates from these teacher models to create a curriculum. We found tentative evidence that our curriculum made it easier for models to acquire linguistic knowledge from the training data: on the subset of tasks in the BabyLM challenge suite evaluating models' grammatical knowledge of English, models first trained on the BabyLM data curriculum and then on a few randomly ordered training epochs performed slightly better than models trained on randomly ordered epochs alone. This improved linguistic knowledge acquisition did not result in better alignment with human reading behavior, however: models trained on the BabyLM dataset (with or without a curriculum) generated predictions that were as misaligned with human behavior as models trained on larger less curated datasets. This suggests that training on developmentally plausible datasets alone is likely insufficient to generate language models capable of accurately predicting human language processing.

📄 PDF Abstract BibTeX arXiv:2311.18761

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Similar Papers 제목 키워드 기반

Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus

2023-01-27 · Alex Warstadt, Leshem Choshen, Aaron Mueller, Adina Williams 외

We present the call for papers for the BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus. This shared task is intended for participants with an interest in small scale language modeling…

Language AcquisitionLanguage ModelingLanguage ModellingNatural Language Understanding

Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

2025-04-10 · Alex Warstadt, Aaron Mueller, Leshem Choshen, Ethan Wilcox 외

Children can acquire language from less than 100 million words of input. Large language models are far less data-efficient: they typically require 3 or 4 orders of magnitude more data and still do not perform as well as …

BabyBabelLM: A Multilingual Benchmark of Developmentally Plausible Training Data

2025-10-11 · Jaap Jumelet, Abdellah Fourtassi, Akari Haga, Bastian Bunzeck 외 arxiv

We present BabyBabelLM, a multilingual collection of datasets modeling the language a person observes from birth until they acquire a native language. We curate developmentally plausible pretraining data aiming to cover …

Do Construction Distributions Shape Formal Language Learning In German BabyLMs?

2025-03-14 · Bastian Bunzeck, Daniel Duran, Sina Zarrieß

We analyze the influence of utterance-level construction distributions in German child-directed speech on the resulting formal linguistic competence and the underlying learning trajectories for small language models trai…

Do Syntactic Categories Help in Developmentally Motivated Curriculum Learning for Language Models?

2025-11-11 · Arzu Burcu Güven, Anna Rogers, Rob van der Goot arxiv

We examine the syntactic properties of BabyLM corpus, and age-groups within CHILDES. While we find that CHILDES does not exhibit strong syntactic differentiation by age, we show that the syntactic knowledge about the tra…