paper-with-me

Papers

VisualScratchpad: Inference-time Visual Concepts Analysis in Vision Language Models

2026-03-07 · Hyesu Lim, Jinho Choi, Taekyung Kim, Byeongho Heo, Jaegul Choo, Dongyoon Han arxiv

High-performing vision language models still produce incorrect answers, yet their failure modes are often difficult to explain. To make model internals more accessible and enable systematic debugging, we introduce VisualScratchpad, an interactive interface for visual concept analysis during inference. We apply sparse autoencoders to the vision encoder and link the resulting visual concepts to text tokens via text-to-image attention, allowing us to examine which visual concepts are both captured by the vision encoder and utilized by the language model. VisualScratchpad also provides a token-latent heatmap view that suggests a sufficient set of latents for effective concept ablation in causal analysis. Through case studies, we reveal three underexplored failure modes: limited cross-modal alignment, misleading visual concepts, and unused hidden cues. Project page: https://hyesulim.github.io/visual_scratchpad_projectpage/

📄 PDF Abstract BibTeX arXiv:2603.07335

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ZeroC: A Neuro-Symbolic Model for Zero-shot Concept Recognition and Acquisition at Inference Time

2022-06-30 · Tailin Wu, Megan Tjandrasuwita, Zhengxuan Wu, Xuelin Yang 외

Humans have the remarkable ability to recognize and acquire novel visual concepts in a zero-shot manner. Given a high-level, symbolic description of a novel concept in terms of previously learned visual concepts and thei…

Novel Concepts

AEGIS: A Mechanism-Guided Defense against Visual Synonym Jailbreaks in Text-to-Image Models

2026-07-07 · Yuanmin Huang, Zhenfei Zhang, Mi Zhang, Geng Hong 외 arxiv

Text-to-image diffusion models have achieved high visual fidelity and broad adoption, but remain vulnerable to safety violations when adversaries exploit them to synthesize illicit content. Existing alignment paradigms, …

CHAIN: Concept-harmonized Hierarchical Inference Interpretation of Deep Convolutional Neural Networks

2020-02-05 · Dan Wang, Xinrui Cui, Z. Jane Wang

With the great success of networks, it witnesses the increasing demand for the interpretation of the internal network mechanism, especially for the net decision-making logic. To tackle the challenge, the Concept-harmoniz…

Decision Making

Text-based inference of moral sentiment change

2020-01-20 · IJCNLP 2019 11 · Jing Yi Xie, Renato Ferreira Pinto Jr., Graeme Hirst, Yang Xu

We present a text-based framework for investigating moral sentiment change of the public via longitudinal corpora. Our framework is based on the premise that language use can inform people's moral perception toward right…

Diachronic Word EmbeddingsWord Embeddings

Explaining generative diffusion models via visual analysis for interpretable decision-making process

2024-02-16 · Ji-Hoon Park, Yeong-Joon Ju, Seong-Whan Lee

Diffusion models have demonstrated remarkable performance in generation tasks. Nevertheless, explaining the diffusion process remains challenging due to it being a sequence of denoising noisy images that are difficult fo…

Decision MakingDenoising