paper-with-me

홈 › Papers

Finding Culture-Sensitive Neurons in Vision-Language Models

2025-10-28 · Xiutian Zhao, Rochelle Choenni, Rohit Saxena, Ivan Titov arxiv

Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs process culturally grounded information, we study the presence of culture-sensitive neurons, i.e., neurons whose activations show preferential sensitivity to inputs associated with particular cultural contexts. We examine whether such neurons are important for culturally diverse visual question answering and where they are located. Using the CVQA benchmark, we identify neurons of culture selectivity and perform diagnostic tests by deactivating the neurons flagged by various identification methods. Experiments on three VLMs across 25 cultural groups demonstrate the existence of neurons whose ablation disproportionately harms performance on questions about the corresponding cultures, while having limited effects on others. Moreover, we introduce a new margin-based selector Contrastive Activation Margin (ConAct) and show that it outperforms probability- and entropy-based methods in identifying neurons associated with cultural selectivity. Finally, our layer-wise analyses reveal that such neurons are not uniformly distributed: they cluster in specific decoder layers in a model-dependent way.

📄 PDF Abstract BibTeX arXiv:2510.24942

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Question Answering

Similar Papers 제목 키워드 기반

Isolating Culture Neurons in Multilingual Large Language Models

2025-08-04 · Danial Namazifard, Lukas Galke Poech arxiv

Language and culture are deeply intertwined, yet it has been unclear how and where multilingual large language models encode culture. Here, we build on an established methodology for identifying language-specific neurons…

Neuron-Level Analysis of Cultural Understanding in Large Language Models

2025-10-09 · Taisei Yamamoto, Ryoma Kumon, Danushka Bollegala, Hitomi Yanaka arxiv

As large language models (LLMs) are increasingly deployed worldwide, ensuring their fair and comprehensive cultural understanding is important. However, LLMs exhibit cultural bias and limited awareness of underrepresente…

Natural Language Understanding

Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation

2025-11-21 · Chuancheng Shi, Shangze Li, Shiming Guo, Simiao Xie 외 arxiv

Multilingual text-to-image (T2I) models have advanced rapidly in terms of visual realism and semantic alignment, and are now widely utilized. Yet outputs vary across cultural contexts: because language carries cultural c…

Text-to-Image Generation

Natural Language Descriptions of Deep Visual Features

2022-01-26 · Evan Hernandez, Sarah Schwettmann, David Bau, Teona Bagashvili 외

Some neurons in deep networks specialize in recognizing highly specific perceptual, structural, or semantic features of inputs. In computer vision, techniques exist for identifying neurons that respond to individual conc…

Attribute

Natural Language Descriptions of Deep Features

2021-09-29 · ICLR 2022 4 · Evan Hernandez, Sarah Schwettmann, David Bau, Teona Bagashvili 외

Some neurons in deep networks specialize in recognizing highly specific perceptual, structural, or semantic features of inputs. In computer vision, techniques exist for identifying neurons that respond to individual conc…

Attribute