paper-with-me

홈 › Papers

The LLM Language Network: A Neuroscientific Approach for Identifying Causally Task-Relevant Units

2024-11-04 · Badr AlKhamissi, Greta Tuckute, Antoine Bosselut, Martin Schrimpf

Large language models (LLMs) exhibit remarkable capabilities on not just language tasks, but also various tasks that are not linguistic in nature, such as logical reasoning and social inference. In the human brain, neuroscience has identified a core language system that selectively and causally supports language processing. We here ask whether similar specialization for language emerges in LLMs. We identify language-selective units within 18 popular LLMs, using the same localization approach that is used in neuroscience. We then establish the causal role of these units by demonstrating that ablating LLM language-selective units -- but not random units -- leads to drastic deficits in language tasks. Correspondingly, language-selective LLM units are more aligned to brain recordings from the human language system than random units. Finally, we investigate whether our localization method extends to other cognitive domains: while we find specialized networks in some LLMs for reasoning and social capabilities, there are substantial differences among models. These findings provide functional and causal evidence for specialization in large language models, and highlight parallels with the functional organization in the brain.

📄 PDF Abstract BibTeX arXiv:2411.02280

Code (1)

bkhmsi/llm-localization 공식 구현 pytorch

Tasks

Logical Reasoning

Similar Papers 제목 키워드 기반

Evaluating Contrast Localizer for Identifying Causal Units in Social & Mathematical Tasks in Language Models

2025-07-31 · Yassine Jamaa, Badr AlKhamissi, Satrajit Ghosh, Martin Schrimpf arxiv

This work adapts a neuroscientific contrast localizer to pinpoint causally relevant units for Theory of Mind (ToM) and mathematical reasoning tasks in large language models (LLMs) and vision-language models (VLMs). Acros…

Mathematical Reasoning

Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs

2026-02-24 · Dhita Putri Pratama, Soyeon Caren Han, Yihao Ding arxiv

Large Vision-Language Models (LVLMs) achieve strong performance on visual question answering benchmarks, yet often rely on spurious correlations rather than genuine causal reasoning. Existing evaluations primarily assess…

Visual Question AnsweringCausal Inference

Brain-inspired probabilistic generative model for double articulation analysis of spoken language

2022-07-06 · Akira Taniguchi, Maoko Muro, Hiroshi Yamakawa, Tadahiro Taniguchi

The human brain, among its several functions, analyzes the double articulation structure in spoken language, i.e., double articulation analysis (DAA). A hierarchical structure in which words are connected to form a sente…

AnatomySentence

Indications of Belief-Guided Agency and Meta-Cognitive Monitoring in Large Language Models

2026-02-02 · Noam Steinmetz Yalon, Ariel Goldstein, Liad Mudrik, Mor Geva arxiv

Rapid advancements in large language models (LLMs) have sparked the question whether these models possess some form of consciousness. To tackle this challenge, Butlin et al. (2023) introduced a list of indicators for con…

Causally Grounded Mechanistic Interpretability for LLMs with Faithful Natural-Language Explanations

2026-02-13 · Ajay Pravin Mahale arxiv

Mechanistic interpretability identifies internal circuits responsible for model behaviors, yet translating these findings into human-understandable explanations remains an open problem. We present a pipeline that bridges…