paper-with-me

홈 › Papers

Can Language Models Induce Grammatical Knowledge from Indirect Evidence?

2024-10-08 · Miyu Oba, Yohei Oseki, Akiyo Fukatsu, Akari Haga, Hiroki Ouchi, Taro Watanabe, Saku Sugawara

What kinds of and how much data is necessary for language models to induce grammatical knowledge to judge sentence acceptability? Recent language models still have much room for improvement in their data efficiency compared to humans. This paper investigates whether language models efficiently use indirect data (indirect evidence), from which they infer sentence acceptability. In contrast, humans use indirect evidence efficiently, which is considered one of the inductive biases contributing to efficient language acquisition. To explore this question, we introduce the Wug InDirect Evidence Test (WIDET), a dataset consisting of training instances inserted into the pre-training data and evaluation instances. We inject synthetic instances with newly coined wug words into pretraining data and explore the model's behavior on evaluation data that assesses grammatical acceptability regarding those words. We prepare the injected instances by varying their levels of indirectness and quantity. Our experiments surprisingly show that language models do not induce grammatical knowledge even after repeated exposure to instances with the same structure but differing only in lexical items from evaluation instances in certain language phenomena. Our findings suggest a potential direction for future research: developing models that use latent indirect evidence to induce grammatical knowledge.

📄 PDF Abstract BibTeX arXiv:2410.06022

Code (0)

등록된 구현이 없습니다.

Tasks

Language AcquisitionSentence

Similar Papers 제목 키워드 기반

Structural Priming Demonstrates Abstract Grammatical Representations in Multilingual Language Models

2023-11-15 · James A. Michaelov, Catherine Arnett, Tyler A. Chang, Benjamin K. Bergen

Abstract grammatical knowledge - of parts of speech and grammatical patterns - is key to the capacity for linguistic generalization in humans. But how abstract is grammatical knowledge in large language models? In the hu…

Sentence

Linear representations of grammaticality in neural language models

2026-07-16 · Jane Li, Najoung Kim arxiv

Whether neural language models (NLMs) possess the ability to distinguish strings on the basis of their grammaticality remains a debated topic in the computational linguistics literature. Existing evidence has largely rel…

Contrast-Space Projection for Network Meta-Analysis: An Exact and Invariant Study-Based Decomposition of Direct and Indirect Contributions

2026-04-23 · Chong Wang, Yanqi Zhang, Zhezhen Jin, Annette O'Connor arxiv

Network meta-analysis (NMA) combines direct and indirect comparisons across a connected treatment network to estimate relative treatment effects. However, there is a lack of exact contribution decompositions that reprodu…

Large Language Model probabilities cannot distinguish between possible and impossible language

2025-09-18 · Evelina Leivada, Raquel Montero, Paolo Morosi, Natalia Moskvina 외 arxiv

A controversial test for Large Language Models concerns the ability to discern possible from impossible language. While some evidence attests to the models' sensitivity to what crosses the limits of grammatically impossi…

Language Models Fail to Introspect About Their Knowledge of Language

2025-03-10 · Siyuan Song, Jennifer Hu, Kyle Mahowald

There has been recent interest in whether large language models (LLMs) can introspect about their own internal states. Such abilities would make LLMs more interpretable, and also validate the use of standard introspectiv…

Sentence