paper-with-me

홈 › Papers

The quasi-semantic competence of LLMs: a case study on the part-whole relation

2025-04-03 · Mattia Proietti, Alessandro Lenci

Understanding the extent and depth of the semantic competence of \emph{Large Language Models} (LLMs) is at the center of the current scientific agenda in Artificial Intelligence (AI) and Computational Linguistics (CL). We contribute to this endeavor by investigating their knowledge of the \emph{part-whole} relation, a.k.a. \emph{meronymy}, which plays a crucial role in lexical organization, but it is significantly understudied. We used data from ConceptNet relations \citep{speer2016conceptnet} and human-generated semantic feature norms \citep{McRae:2005} to explore the abilities of LLMs to deal with \textit{part-whole} relations. We employed several methods based on three levels of analysis: i.) \textbf{behavioral} testing via prompting, where we directly queried the models on their knowledge of meronymy, ii.) sentence \textbf{probability} scoring, where we tested models' abilities to discriminate correct (real) and incorrect (asymmetric counterfactual) \textit{part-whole} relations, and iii.) \textbf{concept representation} analysis in vector space, where we proved the linear organization of the \textit{part-whole} concept in the embedding and unembedding spaces. These analyses present a complex picture that reveals that the LLMs' knowledge of this relation is only partial. They have just a ``\emph{quasi}-semantic'' competence and still fall short of capturing deep inferential properties.

📄 PDF Abstract BibTeX arXiv:2504.02395

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualRelation

Similar Papers 제목 키워드 기반

Competence-Based Analysis of Language Models

2023-03-01 · Adam Davies, Jize Jiang, ChengXiang Zhai

Despite the recent successes of large, pretrained neural language models (LLMs), comparatively little is known about the representations of linguistic structure they learn during pretraining, which can lead to unexpected…

Models Alignment

Emergence and Localisation of Semantic Role Circuits in LLMs

2025-11-25 · Nura Aljaafari, Danilo S. Carvalho, André Freitas arxiv

Despite displaying semantic competence, large language models' internal mechanisms that ground abstract semantic structure remain insufficiently characterised. We propose a method integrating role-cross minimal pairs, te…

Traces of Social Competence in Large Language Models

2026-03-04 · Tom Kouwenhoven, Michiel van der Meer, Max van Duijn arxiv

The False Belief Test (FBT) has been the main method for assessing Theory of Mind (ToM) and related socio-cognitive competencies. For Large Language Models (LLMs), the reliability and explanatory potential of this test h…

Vers un cadre ontologique pour la gestion des comp{é}tences : {à} des fins de formation, de recrutement, de m{é}tier, ou de recherches associ{é}es

2025-07-08 · Ngoc Luyen Le, Marie-Hélène Abel, Bertrand Laforge

The rapid transformation of the labor market, driven by technological advancements and the digital economy, requires continuous competence development and constant adaptation. In this context, traditional competence mana…

Management

Dissociating language and thought in large language models

2023-01-16 · Kyle Mahowald, Anna A. Ivanova, Idan A. Blank, Nancy Kanwisher 외

Large Language Models (LLMs) have come closest among all models to date to mastering human language, yet opinions about their linguistic and cognitive capabilities remain split. Here, we evaluate LLMs using a distinction…