paper-with-me

Papers

Language models align with human judgments on key grammatical constructions

2024-01-19 · Jennifer Hu, Kyle Mahowald, Gary Lupyan, Anna Ivanova, Roger Levy

Do large language models (LLMs) make human-like linguistic generalizations? Dentella et al. (2023) ("DGL") prompt several LLMs ("Is the following sentence grammatically correct in English?") to elicit grammaticality judgments of 80 English sentences, concluding that LLMs demonstrate a "yes-response bias" and a "failure to distinguish grammatical from ungrammatical sentences". We re-evaluate LLM performance using well-established practices and find that DGL's data in fact provide evidence for just how well LLMs capture human behaviors. Models not only achieve high accuracy overall, but also capture fine-grained variation in human linguistic judgments.

📄 PDF Abstract BibTeX arXiv:2402.01676

Code (2)

jennhu/response-to-dgl 공식 구현
jennhu/lm-task-demands

Tasks

Sentence

Similar Papers 제목 키워드 기반

Grammaticality Judgments in Humans and Language Models: Revisiting Generative Grammar with LLMs

2025-12-11 · Lars G. B. Johnsen arxiv

What counts as evidence for syntactic structure? In traditional generative grammar, systematic contrasts in grammaticality such as subject-auxiliary inversion and the licensing of parasitic gaps are taken as evidence for…

A Discerning Several Thousand Judgments: GPT-3 Rates the Article + Adjective + Numeral + Noun Construction

2023-01-29 · Kyle Mahowald

Knowledge of syntax includes knowledge of rare, idiosyncratic constructions. LLMs must overcome frequency biases in order to master such constructions. In this study, I prompt GPT-3 to give acceptability judgments on the…

CoLA

Neural Network Acceptability Judgments

2018-05-31 · TACL 2019 3 · Alex Warstadt, Amanpreet Singh, Samuel R. Bowman

This paper investigates the ability of artificial neural networks to judge the grammatical acceptability of a sentence, with the goal of testing their linguistic competence. We introduce the Corpus of Linguistic Acceptab…

CoLAGeneral ClassificationLanguage AcquisitionLinguistic Acceptability+1

Grammaticality Representation in ChatGPT as Compared to Linguists and Laypeople

2024-06-17 · Zhuang Qiu, Xufeng Duan, Zhenguang G. Cai

Large language models (LLMs) have demonstrated exceptional performance across various linguistic tasks. However, it remains uncertain whether LLMs have developed human-like fine-grained grammatical intuition. This prereg…

AttributeSentence

Investigating representations of verb bias in neural language models

2020-10-05 · EMNLP 2020 11 · Robert D. Hawkins, Takateru Yamakoshi, Thomas L. Griffiths, Adele E. Goldberg

Languages typically provide more than one grammatical construction to express certain types of messages. A speaker's choice of construction is known to depend on multiple factors, including the choice of main verb -- a p…

Sentence