paper-with-me

홈 › Papers

SCoCCA: Multi-modal Sparse Concept Decomposition via Canonical Correlation Analysis

2026-03-14 · Ehud Gordon, Meir Yossef Levi, Guy Gilboa arxiv

Interpreting the internal reasoning of vision-language models is essential for deploying AI in safety-critical domains. Concept-based explainability provides a human-aligned lens by representing a model's behavior through semantically meaningful components. However, existing methods are largely restricted to images and overlook the cross-modal interactions. Text-image embeddings, such as those produced by CLIP, suffer from a modality gap, where visual and textual features follow distinct distributions, limiting interpretability. Canonical Correlation Analysis (CCA) offers a principled way to align features from different distributions, but has not been leveraged for multi-modal concept-level analysis. We show that the objectives of CCA and InfoNCE are closely related, such that optimizing CCA implicitly optimizes InfoNCE, providing a simple, training-free mechanism to enhance cross-modal alignment without affecting the pre-trained InfoNCE objective. Motivated by this observation, we couple concept-based explainability with CCA, introducing Concept CCA (CoCCA), a framework that aligns cross-modal embeddings while enabling interpretable concept decomposition. We further extend it and propose Sparse Concept CCA (SCoCCA), which enforces sparsity to produce more disentangled and discriminative concepts, facilitating improved activation, ablation, and semantic manipulation. Our approach generalizes concept-based explanations to multi-modal embeddings and achieves state-of-the-art performance in concept discovery, evidenced by reconstruction and manipulation tasks such as concept ablation.

📄 PDF Abstract BibTeX arXiv:2603.13884

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Decomposing multimodal embedding spaces with group-sparse autoencoders

2026-01-27 · Chiraag Kaushik, Davis Barch, Andrea Fanelli arxiv

The Linear Representation Hypothesis asserts that the embeddings learned by neural networks can be understood as linear combinations of features corresponding to high-level concepts. Based on this ansatz, sparse autoenco…

ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition

2026-05-14 · Shen Lin, Jing Lin, Junhao Dong, Piotr Koniusz 외 arxiv

Machine unlearning in Vision-Language Models (VLMs) is typically performed at the image or instance level, making it difficult to precisely remove target knowledge without affecting unrelated semantics. This issue is esp…

Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE)

2024-02-16 · Usha Bhalla, Alex Oesterling, Suraj Srinivas, Flavio P. Calmon 외

CLIP embeddings have demonstrated remarkable performance across a wide range of multimodal applications. However, these high-dimensional, dense vector representations are not easily interpretable, limiting our understand…

Model Editing

Concepts from Representations: Post-hoc Concept Bottleneck Models via Sparse Decomposition of Visual Representations

2026-01-18 · Shizhan Gong, Xiaofan Zhang, Qi Dou arxiv

Deep learning has achieved remarkable success in image recognition, yet their inherent opacity poses challenges for deployment in critical domains. Concept-based interpretations aim to address this by explaining model re…

Image Classification

Coupled generator decomposition for fusion of electro- and magnetoencephalography data

2024-03-02 · Anders Stevnhoved Olsen, Jesper Duemose Nielsen, Morten Mørup

Data fusion modeling can identify common features across diverse data sources while accounting for source-specific variability. Here we introduce the concept of a \textit{coupled generator decomposition} and demonstrate …

EEGStochastic Optimization