paper-with-me

홈 › Papers

MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model

2024-06-17 · Jiahao Huo, Yibo Yan, Boren Hu, Yutao Yue, Xuming Hu

Projecting visual features into word embedding space has become a significant fusion strategy adopted by Multimodal Large Language Models (MLLMs). However, its internal mechanisms have yet to be explored. Inspired by multilingual research, we identify domain-specific neurons in multimodal large language models. Specifically, we investigate the distribution of domain-specific neurons and the mechanism of how MLLMs process features from diverse domains. Furthermore, we propose a three-stage mechanism for language model modules in MLLMs when handling projected image features, and verify this hypothesis using logit lens. Extensive experiments indicate that while current MLLMs exhibit Visual Question Answering (VQA) capability, they may not fully utilize domain-specific information. Manipulating domain-specific neurons properly will result in a 10% change of accuracy at most, shedding light on the development of cross-domain, all-encompassing MLLMs in the future. The source code is available at https://github.com/Z1zs/MMNeuron.

📄 PDF Abstract BibTeX arXiv:2406.11193

Code (1)

z1zs/mmneuron 공식 구현 pytorch

Tasks

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language ModelQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

Choose Your Neuron: Incorporating Domain Knowledge through Neuron-Importance

2018-08-08 · ECCV 2018 9 · Ramprasaath R. Selvaraju, Prithvijit Chattopadhyay, Mohamed Elhoseiny, Tilak Sharma 외

Individual neurons in convolutional neural networks supervised for image-level classification tasks have been shown to implicitly learn semantically meaningful concepts ranging from simple textures and shapes to whole or…

Generalized Zero-Shot LearningZero-Shot Learning

Discovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models

2026-01-06 · Xiutian Zhao, Björn Schuller, Berrak Sisman arxiv

Emotion is a central dimension of spoken communication, yet, we still lack a mechanistic account of how modern large audio-language models (LALMs) encode it internally. We present the first neuron-level interpretability …

Emotion Recognition

Neuron Level Analysis of Large Language Model in Legal Domain Reasoning

2026-06-14 · Eri Onami, Youmi Ma, Shuhei Kurita, Naoaki Okazaki arxiv

We presented a neuron-level analysis of legal-domain reasoning in LLMs, comparing it with other applied domain tasks across seven open-weight models. Using neuron attribution scores to rank and suppress influential neuro…

Discovering Influential Neuron Path in Vision Transformers

2025-03-12 · Yifan Wang, Yifei Liu, Yingdong Shi, Changming Li 외

Vision Transformer models exhibit immense power yet remain opaque to human understanding, posing challenges and risks for practical applications. While prior research has attempted to demystify these models through input…

image-classificationImage Classification

HINT: Hierarchical Neuron Concept Explainer

2022-03-27 · CVPR 2022 1 · Andong Wang, Wei-Ning Lee, Xiaojuan Qi

To interpret deep networks, one main approach is to associate neurons with human-understandable concepts. However, existing methods often ignore the inherent relationships of different concepts (e.g., dog and cat both be…

Object LocalizationWeakly-Supervised Object Localization