paper-with-me

홈 › Papers

Discovering Decoupled Functional Modules in Large Language Models

2026-03-18 · Yanke Yu, Jin Li, Ying Sun, Ping Li, Zhefeng Wang, Yi Zheng arxiv

Understanding the internal functional organization of Large Language Models (LLMs) is crucial for improving their trustworthiness and performance. However, how LLMs organize different functions into modules remains highly unexplored. To bridge this gap, we formulate a functional module discovery problem and propose an Unsupervised LLM Cross-layer MOdule Discovery (ULCMOD) framework that simultaneously disentangles the large set of neurons in the entire LLM into modules while discovering the topics of input samples related to these modules. Our framework introduces a novel objective function and an efficient Iterative Decoupling (IterD) algorithm. Extensive experiments show that our method discovers high-quality, disentangled modules that capture more meaningful semantic information and achieve superior performance in various downstream tasks. Moreover, our qualitative analysis reveals that the discovered modules show semantic coherence, correspond to interpretable specializations, and a clear spatial and hierarchical organization within the LLM. Our work provides a novel tool for interpreting the functional modules of LLMs, filling a critical blank in LLM's interpretability research.

📄 PDF Abstract BibTeX arXiv:2603.17823

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Discovering Dynamic Functional Brain Networks via Spatial and Channel-wise Attention

2022-05-19 · Yiheng Liu, Enjie Ge, Mengshen He, Zhengliang Liu 외

Using deep learning models to recognize functional brain networks (FBNs) in functional magnetic resonance imaging (fMRI) has been attracting increasing interest recently. However, most existing work focuses on detecting …

Functional Connectivity

Decoupled Functional Evaluation of Autonomous Driving Models via Feature Map Quality Scoring

2025-08-11 · Ludan Zhang, Sihan Wang, Yuqi Dai, Shuofei Qiao 외 arxiv

End-to-end models are emerging as the mainstream in autonomous driving perception and planning. However, the lack of explicit supervision signals for intermediate functional modules leads to opaque operational mechanisms…

3D Object DetectionAutonomous Driving

VisualSimpleQA: A Benchmark for Decoupled Evaluation of Large Vision-Language Models in Fact-Seeking Question Answering

2025-03-09 · Yanling Wang, Yihan Zhao, Xiaodong Chen, Shasha Guo 외

Large vision-language models (LVLMs) have demonstrated remarkable achievements, yet the generation of non-factual responses remains prevalent in fact-seeking question answering (QA). Current multimodal fact-seeking bench…

Question Answering

CBDES MoE: Hierarchically Decoupled Mixture-of-Experts for Functional Modules in Autonomous Driving

2025-08-11 · Qi Xiang, Kunsong Shi, Zhigui Lin, Lei He arxiv

Bird's Eye View (BEV) perception systems based on multi-sensor feature fusion have become a fundamental cornerstone for end-to-end autonomous driving. However, existing multi-modal BEV methods commonly suffer from limite…

3D Object DetectionAutonomous Driving

Generative artificial intelligence-enabled dynamic detection of nicotine-related circuits

2022-12-13 · Changwei Gong, Changhong Jing, Ye Li, Xinan Liu 외

The identification of addiction-related circuits is critical for explaining addiction processes and developing addiction treatments. And models of functional addiction circuits developed from functional imaging are an ef…

Contrastive Learning