paper-with-me

홈 › Papers

Unveiling A Core Linguistic Region in Large Language Models

2023-10-23 · Jun Zhao, Zhihao Zhang, Yide Ma, Qi Zhang, Tao Gui, Luhui Gao, Xuanjing Huang

Brain localization, which describes the association between specific regions of the brain and their corresponding functions, is widely accepted in the field of cognitive science as an objective fact. Today's large language models (LLMs) possess human-level linguistic competence and can execute complex tasks requiring abstract knowledge and reasoning. To deeply understand the inherent mechanisms of intelligence emergence in LLMs, this paper conducts an analogical research using brain localization as a prototype. We have discovered a core region in LLMs that corresponds to linguistic competence, accounting for approximately 1% of the total model parameters. This core region exhibits significant dimension dependency, and perturbations to even a single parameter on specific dimensions can lead to a loss of linguistic competence. Furthermore, we observe that an improvement in linguistic competence does not necessarily accompany an elevation in the model's knowledge level, which might imply the existence of regions of domain knowledge that are dissociated from the linguistic region. Overall, exploring the LLMs' functional regions provides insights into the foundation of their intelligence. In the future, we will continue to investigate knowledge regions within LLMs and the interactions between them.

📄 PDF Abstract BibTeX arXiv:2310.14928

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unveiling Linguistic Regions in Large Language Models

2024-02-22 · Zhihao Zhang, Jun Zhao, Qi Zhang, Tao Gui 외

Large Language Models (LLMs) have demonstrated considerable cross-lingual alignment and generalization ability. Current research primarily focuses on improving LLMs' cross-lingual generalization capabilities. However, th…

UNVEILING: What Makes Linguistics Olympiad Puzzles Tricky for LLMs?

2025-08-15 · Mukund Choudhary, KV Aditya Srivatsa, Gaurja Aeron, Antara Raaghavi Bhattacharya 외 arxiv

Large language models (LLMs) have demonstrated potential in reasoning tasks, but their performance on linguistics puzzles remains consistently poor. These puzzles, often derived from Linguistics Olympiad (LO) contests, p…

Keynote: Unveiling the Linguistic Weaknesses of Neural MT

2018-03-01 · WS 2018 3 · Arianna Bisazza
Machine Translation

Unveiling Language Competence Neurons: A Psycholinguistic Approach to Model Interpretability

2024-09-24 · Xufeng Duan, Xinyu Zhou, Bei Xiao, Zhenguang G. Cai

As large language models (LLMs) advance in their linguistic capacity, understanding how they capture aspects of language competence remains a significant challenge. This study therefore employs psycholinguistic paradigms…

Language ModelingLanguage Modelling

Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing

2025-02-21 · Zhilin Wang, Yafu Li, Jianhao Yan, Yu Cheng 외

Dynamical systems theory provides a framework for analyzing iterative processes and evolution over time. Within such systems, repetitive transformations can lead to stable configurations, known as attractors, including f…

Diversity