paper-with-me

Papers

LLMs Learn Constructions That Humans Do Not Know

2025-08-22 · Jonathan Dunn, Mai Mohamed Eida arxiv

This paper investigates false positive constructions: grammatical structures which an LLM hallucinates as distinct constructions but which human introspection does not support. Both a behavioural probing task using contextual embeddings and a meta-linguistic probing task using prompts are included, allowing us to distinguish between implicit and explicit linguistic knowledge. Both methods reveal that models do indeed hallucinate constructions. We then simulate hypothesis testing to determine what would have happened if a linguist had falsely hypothesized that these hallucinated constructions do exist. The high accuracy obtained shows that such false hypotheses would have been overwhelmingly confirmed. This suggests that construction probing methods suffer from a confirmation bias and raises the issue of what unknown and incorrect syntactic knowledge these models also possess.

📄 PDF Abstract BibTeX arXiv:2508.16837

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When the LM misunderstood the human chuckled: Analyzing garden path effects in humans and language models

2025-02-13 · Samuel Joseph Amouyal, Aya Meltzer-Asscher, Jonathan Berant

Modern Large Language Models (LLMs) have shown human-like abilities in many language tasks, sparking interest in comparing LLMs' and humans' language processing. In this paper, we conduct a detailed comparison of the two…

Image GenerationSentenceText to Image GenerationText-to-Image Generation

Grammaticality Judgments in Humans and Language Models: Revisiting Generative Grammar with LLMs

2025-12-11 · Lars G. B. Johnsen arxiv

What counts as evidence for syntactic structure? In traditional generative grammar, systematic contrasts in grammaticality such as subject-auxiliary inversion and the licensing of parasitic gaps are taken as evidence for…

Assessing Language Comprehension in Large Language Models Using Construction Grammar

2025-01-08 · Wesley Scivetti, Melissa Torgbi, Austin Blodgett, Mollie Shichman 외

Large Language Models, despite their significant capabilities, are known to fail in surprising and unpredictable ways. Evaluating their true `understanding' of language is particularly challenging due to the extensive we…

Natural Language InferenceNatural Language Understanding

Evaluating LLMs on Chinese Topic Constructions: A Research Proposal Inspired by Tian et al. (2024)

2025-04-21 · Xiaodong Yang

This paper proposes a framework for evaluating large language models (LLMs) on Chinese topic constructions, focusing on their sensitivity to island constraints. Drawing inspiration from Tian et al. (2024), we outline an …

Experimental DesignSensitivity

LLMs grasp morality in concept

2023-11-04 · Mark Pock, Andre Ye, Jared Moore

Work in AI ethics and fairness has made much progress in regulating LLMs to reflect certain values, such as fairness, truth, and diversity. However, it has taken the problem of how LLMs might 'mean' anything at all for g…

DiversityEthicsFairnessPhilosophy+1