paper-with-me

홈 › Papers

Generics are puzzling. Can language models find the missing piece?

2024-12-15 · Gustavo Cilleruelo Calderón, Emily Allaway, Barry Haddow, Alexandra Birch

Generic sentences express generalisations about the world without explicit quantification. Although generics are central to everyday communication, building a precise semantic framework has proven difficult, in part because speakers use generics to generalise properties with widely different statistical prevalence. In this work, we study the implicit quantification and context-sensitivity of generics by leveraging language models as models of language. We create ConGen, a dataset of 2873 naturally occurring generic and quantified sentences in context, and define p-acceptability, a metric based on surprisal that is sensitive to quantification. Our experiments show generics are more context-sensitive than determiner quantifiers and about 20% of naturally occurring generics we analyze express weak generalisations. We also explore how human biases in stereotypes can be observed in language models.

📄 PDF Abstract BibTeX arXiv:2412.11318

Code (1)

ilyocoris/generics_are_puzzling 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Generics in science communication: Misaligned interpretations across laypeople, scientists, and large language models

2026-02-05 · Uwe Peters, Andrea Bertazzoli, Jasmine M. DeJesus, Gisela J. van der Velden 외 arxiv

Scientists often use generics, that is, unquantified statements about whole categories of people or phenomena, when communicating research findings (e.g., "statins reduce cardiovascular events"). Large language models (L…

GenericsKB: A Knowledge Base of Generic Statements

2020-05-02 · Sumithra Bhakthavatsalam, Chloe Anastasiades, Peter Clark

We present a new resource for the NLP community, namely a large (3.5M+ sentence) knowledge base of *generic statements*, e.g., "Trees remove carbon dioxide from the atmosphere", collected from multiple corpora. This is t…

Sentence

Generics and Default Reasoning in Large Language Models

2025-08-19 · James Ravi Kirkpatrick, Rachel Katharine Sterken arxiv

This paper evaluates the capabilities of 28 large language models (LLMs) to reason with 20 defeasible reasoning patterns involving generic generalizations (e.g., 'Birds fly', 'Ravens are black') central to non-monotonic …

Generic Overgeneralization in Pre-trained Language Models

2022-10-01 · COLING 2022 10 · Sello Ralethe, Jan Buys

Generic statements such as “ducks lay eggs” make claims about kinds, e.g., ducks as a category. The generic overgeneralization effect refers to the inclination to accept false universal generalizations such as “all ducks…

Are Generics and Negativity about Social Groups Common on Social Media? A Comparative Analysis of Twitter (X) Data

2024-05-14 · Uwe Peters, Ignacio Ojea Quintana

Generics (unquantified generalizations) are thought to be pervasive in communication and when they are about social groups, this may offend and polarize people because generics gloss over variations between individuals. …