paper-with-me

홈 › Papers

Traces of Social Competence in Large Language Models

2026-03-04 · Tom Kouwenhoven, Michiel van der Meer, Max van Duijn arxiv

The False Belief Test (FBT) has been the main method for assessing Theory of Mind (ToM) and related socio-cognitive competencies. For Large Language Models (LLMs), the reliability and explanatory potential of this test have remained limited due to issues like data contamination, insufficient model details, and inconsistent controls. We address these issues by testing 17 open-weight models on a balanced set of 192 FBT variants (Trott et al., 2023) using Bayesian Logistic regression to identify how model size and post-training affect socio-cognitive competence. We find that scaling model size benefits performance, but not strictly. A cross-over effect reveals that explicating propositional attitudes (X thinks) fundamentally alters response patterns. Instruction tuning partially mitigates this effect, but further reasoning-oriented fine-tuning amplifies it. In a case study analysing social reasoning ability throughout OLMo 2 training, we show that this cross-over effect emerges during pre-training, suggesting that models acquire stereotypical response patterns tied to mental-state vocabulary that can outweigh other scenario semantics. Finally, vector steering allows us to isolate a think vector as the causal driver of observed FBT behaviour.

📄 PDF Abstract BibTeX arXiv:2603.04161

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Annotating Dimensions of Social Perception in Text: A Sentence-Level Dataset of Warmth and Competence

2026-01-09 · Mutaz Ayesh, Saif M. Mohammad, Nedjma Ousidhoum arxiv

Warmth (W) (often further broken down intoTrust (T) and Sociability (S)) and Competence (C) are central dimensions along which people evaluate individuals and social groups (Fiske, 2018). While these constructs are well …

Do LLM Agents Know How to Ground, Recover, and Assess? A Benchmark for Epistemic Competence in Information-Seeking Agents

2025-09-26 · Jiaqi Shao, Yuxiang Lin, Munish Prasad Lohani, Yufeng Miao 외 arxiv

Recent work has explored training Large Language Model (LLM) search agents with reinforcement learning (RL) for open-domain question answering (QA). However, most evaluations focus solely on final answer accuracy, overlo…

Open-Domain Question AnsweringReinforcement Learning

How Annotation Trains Annotators: Competence Development in Social Influence Recognition

2026-04-03 · Maciej Markiewicz, Beata Bajcar, Wiktoria Mieleszczenko-Kowszewicz, Aleksander Szczęsny 외 arxiv

Human data annotation, especially when involving experts, is often treated as an objective reference. However, many annotation tasks are inherently subjective, and annotators' judgments may evolve over time. This study i…

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction

2026-03-16 · Tanvir Ahmed Sijan, S. M Golam Rifat, Pankaj Chowdhury Partha, Md. Tanjeed Islam 외 arxiv

Large Language Models have demonstrated strong multilingual fluency, yet fluency alone does not guarantee socially appropriate language use. In high-context languages, communicative competence requires sensitivity to soc…

StereoMap: Quantifying the Awareness of Human-like Stereotypes in Large Language Models

2023-10-20 · Sullam Jeoung, Yubin Ge, Jana Diesner

Large Language Models (LLMs) have been observed to encode and perpetuate harmful associations present in the training data. We propose a theoretically grounded framework called StereoMap to gain insights into their perce…