paper-with-me

홈 › Papers

Head Pursuit: Probing Attention Specialization in Multimodal Transformers

2025-10-24 · Lorenzo Basile, Valentino Maiorca, Diego Doimo, Francesco Locatello, Alberto Cazzaniga arxiv

Language and vision-language models have shown impressive performance across a wide range of tasks, but their internal mechanisms remain only partly understood. In this work, we study how individual attention heads in text-generative models specialize in specific semantic or visual attributes. Building on an established interpretability method, we reinterpret the practice of probing intermediate activations with the final decoding layer through the lens of signal processing. This lets us analyze multiple samples in a principled way and rank attention heads based on their relevance to target concepts. Our results show consistent patterns of specialization at the head level across both unimodal and multimodal transformers. Remarkably, we find that editing as few as 1% of the heads, selected using our method, can reliably suppress or enhance targeted concepts in the model output. We validate our approach on language tasks such as question answering and toxicity mitigation, as well as vision-language tasks including image classification and captioning. Our findings highlight an interpretable and controllable structure within attention layers, offering simple tools for understanding and editing large-scale generative models.

📄 PDF Abstract BibTeX arXiv:2510.21518

Code (0)

등록된 구현이 없습니다.

Tasks

Image ClassificationQuestion Answering

Similar Papers 제목 키워드 기반

Semantic Head Specialization Guides Hybrid ViT Attention for Multimodal LLMs

2026-08-28 · Chenhong He, Lei Li, Shicheng Li, Hanglong Lv 외 arxiv

Hybrid attention dominates frontier LLMs, yet Vision Transformers (ViTs) in multimodal LLMs lack a satisfactory hybrid design, with no consensus on why certain attention patterns work better. To fill this gap, we study V…

Head-wise Modality Specialization within MLLMs for Robust Fake News Detection under Missing Modality

2026-04-08 · Kai Qian, Weijie Shi, Jiaqi Wang, Mengze Li 외 arxiv

Multimodal fake news detection (MFND) aims to verify news credibility by jointly exploiting textual and visual evidence. However, real-world news dissemination frequently suffers from missing modality due to deleted imag…

Fake News Detection

Behind the Scene: Revealing the Secrets of Pre-trained Vision-and-Language Models

2020-05-15 · ECCV 2020 8 · Jize Cao, Zhe Gan, Yu Cheng, Licheng Yu 외

Recent Transformer-based large-scale pre-trained models have revolutionized vision-and-language (V+L) research. Models such as ViLBERT, LXMERT and UNITER have significantly lifted state of the art across a wide range of …

coreference-resolutionCoreference Resolutioncross-modal alignment

Interpreting and Exploiting Functional Specialization in Multi-Head Attention under Multi-task Learning

2023-10-16 · Chong Li, Shaonan Wang, Yunhao Zhang, Jiajun Zhang 외

Transformer-based models, even though achieving super-human performance on several downstream tasks, are often regarded as a black box and used as a whole. It is still unclear what mechanisms they have learned, especiall…

Multi-Task LearningTransfer Learning

Specialization of softmax attention heads: insights from the high-dimensional single-location model

2026-03-04 · M. Sagitova, O. Duranthon, L. Zdeborová arxiv

Multi-head attention enables transformer models to represent multiple attention patterns simultaneously. Empirically, head specialization emerges in distinct stages during training, while many heads remain redundant and …