paper-with-me

Papers

Interactivity x Explainability: Toward Understanding How Interactivity Can Improve Computer Vision Explanations

2025-04-14 · Indu Panigrahi, Sunnie S. Y. Kim, Amna Liaqat, Rohan Jinturkar, Olga Russakovsky, Ruth Fong, Parastoo Abtahi

Explanations for computer vision models are important tools for interpreting how the underlying models work. However, they are often presented in static formats, which pose challenges for users, including information overload, a gap between semantic and pixel-level information, and limited opportunities for exploration. We investigate interactivity as a mechanism for tackling these issues in three common explanation types: heatmap-based, concept-based, and prototype-based explanations. We conducted a study (N=24), using a bird identification task, involving participants with diverse technical and domain expertise. We found that while interactivity enhances user control, facilitates rapid convergence to relevant information, and allows users to expand their understanding of the model and explanation, it also introduces new challenges. To address these, we provide design recommendations for interactive computer vision explanations, including carefully selected default views, independent input controls, and constrained output spaces.

📄 PDF Abstract BibTeX arXiv:2504.10745

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

2026-06-09 · Atsumoto Ohashi, Neil Zeghidour, Alexandre Défossez, Eugene Kharitonov arxiv

Full-duplex spoken dialogue models can listen and speak simultaneously, making them a promising architecture for natural conversation. However, current models are trained solely with supervised learning through token-lev…

Reinforcement LearningDialogue Evaluation

Explainability Requires Interactivity

2021-09-16 · Matthias Kirchler, Martin Graf, Marius Kloft, Christoph Lippert

When explaining the decisions of deep neural networks, simple stories are tempting but dangerous. Especially in computer vision, the most popular explanation approaches give a false sense of comprehension to its users an…

Vocal Interactivity in Crowds, Flocks and Swarms: Implications for Voice User Interfaces

2019-07-26 · Roger K. Moore

Recent years have seen an explosion in the availability of Voice User Interfaces. However, user surveys suggest that there are issues with respect to usability, and it has been hypothesised that contemporary voice-enable…

The World Is Bigger! A Computationally-Embedded Perspective on the Big World Hypothesis

2025-12-29 · Alex Lewandowski, Adtiya A. Ramesh, Edan Meyer, Dale Schuurmans 외 arxiv

Continual learning is often motivated by the idea, known as the big world hypothesis, that "the world is bigger" than the agent. Recent problem formulations capture this idea by explicitly constraining an agent relative …

Reinforcement LearningContinual Learning

AnyTalker: Scaling Multi-Person Talking Video Generation with Interactivity Refinement

2025-11-28 · Zhizhou Zhong, Yicheng Ji, Zhe Kong, Yiying Liu 외 arxiv

Recently, multi-person video generation has started to gain prominence. While a few preliminary works have explored audio-driven multi-person talking video generation, they often face challenges due to the high costs of …

Video Generation