paper-with-me

홈 › Papers

Do Large Language Models know who did what to whom?

2025-04-23 · Joseph M. Denning, Xiaohan, Guo, Bryor Snefjella, Idan A. Blank

Large Language Models (LLMs) are commonly criticized for not understanding language. However, many critiques focus on cognitive abilities that, in humans, are distinct from language processing. Here, we instead study a kind of understanding tightly linked to language: inferring who did what to whom (thematic roles) in a sentence. Does the central training objective of LLMs-word prediction-result in sentence representations that capture thematic roles? In two experiments, we characterized sentence representations in four LLMs. In contrast to human similarity judgments, in LLMs the overall representational similarity of sentence pairs reflected syntactic similarity but not whether their agent and patient assignments were identical vs. reversed. Furthermore, we found little evidence that thematic role information was available in any subset of hidden units. However, some attention heads robustly captured thematic roles, independently of syntax. Therefore, LLMs can extract thematic roles but, relative to humans, this information influences their representations more weakly.

📄 PDF Abstract BibTeX arXiv:2504.16884

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Focus 설명 없음

Similar Papers 제목 키워드 기반

The Five Ws of Multi-Agent Communication: Who Talks to Whom, When, What, and Why -- A Survey from MARL to Emergent Language and LLMs

2026-02-12 · Jingdi Chen, Hanqing Yang, Zongjun Liu, Carlee Joe-Wong arxiv

Multi-agent sequential decision-making powers many real-world systems, from autonomous vehicles and robotics to collaborative AI assistants. In dynamic, partially observable environments, communication is often what redu…

Multi-agent Reinforcement LearningAutonomous Vehicles

Who Did What to Whom? A Contrastive Study of Syntacto-Semantic Dependencies

2012-07-01 · WS 2012 7 · Angelina Ivanova, Stephan Oepen, Lilja {\O}vrelid, Dan Flickinger
Dependency ParsingMachine TranslationSentiment Analysis

Who Sides with Whom? Towards Computational Construction of Discourse Networks for Political Debates

2019-07-01 · ACL 2019 7 · Sebastian Pad{\'o}, Andre Blessing, Nico Blokker, Erenay Dayanik 외

Understanding the structures of political debates (which actors make what claims) is essential for understanding democratic political decision making. The vision of computational construction of such discourse networks f…

Decision MakingKnowledge Base Population

A Psycholinguistic Evaluation of Language Models' Sensitivity to Argument Roles

2024-10-21 · Eun-Kyoung Rosa Lee, Sathvik Nair, Naomi Feldman

We present a systematic evaluation of large language models' sensitivity to argument roles, i.e., who did what to whom, by replicating psycholinguistic studies on human argument role processing. In three experiments, we …

SensitivitySentence

A Classification of Artificial Intelligence Systems for Mathematics Education

2021-07-13 · Steven Van Vaerenbergh, Adrián Pérez-Suay

This chapter provides an overview of the different Artificial Intelligence (AI) systems that are being used in contemporary digital tools for Mathematics Education (ME). It is aimed at researchers in AI and Machine Learn…

Classification