paper-with-me

홈 › Papers

Metonymy in vision models undermines attention-based interpretability

2026-05-07 · Ananthu Aniraj, Cassio F. Dantas, Dino Ienco, Massimiliano Mancini, Diego Marcos arxiv

Part-based reasoning is a classical strategy to make a computer vision model directly focus on the object parts that are relevant to the downstream task. In the context of deep learning, this also serves to improve by-design interpretability, often by using part-centric attention mechanisms on top of a latent image representation provided by a standard, black-box model. This approach is based on a locality assumption: that the latent representation of an object part encodes primarily information about the corresponding image region. In this work, we test this basic assumption, measuring intra-object leakage in vision models using part-based attribute annotations. Through a comprehensive experimental evaluation, we show that modern pretrained vision transformers violate the locality assumption and exhibit a strong intra-object leakage, in which each part encodes information from the whole object, a visual metonymy that compromises the faithfulness of attention-based interpretable-by-design methods for part-based reasoning, ultimately rendering them uninterpretable. In addition, we establish an upper bound using a two-stage approach that prevents leakage by design. We then show that this inherently disentangled feature extraction improves attribute-driven part discovery on a variety of tasks, confirming the practical impact of intra-object leakage. Our results uncover a neglected issue affecting the interpretability of part-based representations, such as those in CBMs relying on part-centric concepts, highlighting that two-stage approaches offer a promising way to mitigate it.

📄 PDF Abstract BibTeX arXiv:2605.06095

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Paraphrasing verbal metonymy through computational methods

2017-09-18 · Alberto Morón Hernández

Verbal metonymy has received relatively scarce attention in the field of computational linguistics despite the fact that a model to accurately paraphrase metonymy has applications both in academia and the technology sect…

Impact of Target Word and Context on End-to-End Metonymy Detection

2021-12-06 · Kevin Alex Mathews, Michael Strube

Metonymy is a figure of speech in which an entity is referred to by another related entity. The task of metonymy detection aims to distinguish metonymic tokens from literal ones. Until now, metonymy detection methods att…

Sentence

How Universal is Metonymy? Results from a Large-Scale Multilingual Analysis

2022-07-01 · NAACL (SIGTYP) 2022 7 · Temuulen Khishigsuren, Gábor Bella, Thomas Brochhagen, Daariimaa Marav 외

Metonymy is regarded by most linguists as a universal cognitive phenomenon, especially since the emergence of the theory of conceptual mappings. However, the field data backing up claims of universality has not been larg…

A Computational Approach to Visual Metonymy

2026-01-25 · Saptarshi Ghosh, Linfeng Liu, Tianyu Jiang arxiv

Images often communicate more than they literally depict: a set of tools can suggest an occupation and a cultural artifact can suggest a tradition. This kind of indirect visual reference, known as visual metonymy, invite…

A Large Harvested Corpus of Location Metonymy

2020-05-01 · LREC 2020 5 · Kevin Alex Mathews, Michael Strube

Metonymy is a figure of speech in which an entity is referred to by another related entity. The existing datasets of metonymy are either too small in size or lack sufficient coverage. We propose a new, labelled, high-qua…