paper-with-me

Papers

Do language models learn typicality judgments from text?

2021-05-06 · Kanishka Misra, Allyson Ettinger, Julia Taylor Rayz

Building on research arguing for the possibility of conceptual and categorical knowledge acquisition through statistics contained in language, we evaluate predictive language models (LMs) -- informed solely by textual input -- on a prevalent phenomenon in cognitive science: typicality. Inspired by experiments that involve language processing and show robust typicality effects in humans, we propose two tests for LMs. Our first test targets whether typicality modulates LM probabilities in assigning taxonomic category memberships to items. The second test investigates sensitivities to typicality in LMs' probabilities when extending new information about items to their categories. Both tests show modest -- but not completely absent -- correspondence between LMs and humans, suggesting that text-based exposure alone is insufficient to acquire typicality knowledge.

📄 PDF Abstract BibTeX arXiv:2105.02987

Code (1)

kanishkamisra/typicalityprobing 공식 구현 pytorch

Similar Papers 제목 키워드 기반

How Well Do Deep Learning Models Capture Human Concepts? The Case of the Typicality Effect

2024-05-25 · Siddhartha K. Vemuri, Raj Sanjay Shah, Sashank Varma

How well do representations learned by ML models align with those of humans? Here, we consider concept representations learned by deep learning models and evaluate whether they show a fundamental behavioral signature of …

Language ModelingLanguage Modelling

Graded Causation and Defaults

2013-09-05 · Joseph Y. Halpern, Christopher Hitchcock

Recent work in psychology and experimental philosophy has shown that judgments of actual causation are often influenced by consideration of defaults, typicality, and normality. A number of philosophers and computer scien…

Philosophy

Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics

2026-01-08 · Subhadeep Roy, Gagan Bhatia, Steffen Eger arxiv

Automatic metrics are widely used to evaluate text-to-image models, often replacing human judgment in benchmarking, model selection, and large-scale data filtering. Yet they may reward images that look plausible or proto…

Revealing interpretable object representations from human behavior

2019-01-09 · ICLR 2019 5 · Charles Y. Zheng, Francisco Pereira, Chris I. Baker, Martin N. Hebart

To study how mental object representations are related to behavior, we estimated sparse, non-negative representations of objects using human behavioral judgments on images representative of 1,854 object categories. These…

Object

Benchmarking VLMs' Reasoning About Persuasive Atypical Images

2024-09-16 · Sina Malakouti, Aysan Aghazadeh, Ashmit Khandelwal, Adriana Kovashka

Vision language models (VLMs) have shown strong zero-shot generalization across various tasks, especially when integrated with large language models (LLMs). However, their ability to comprehend rhetorical and persuasive …

BenchmarkingObject RecognitionZero-shot Generalization