paper-with-me

홈 › Papers

Does Explanation Correctness Matter? Linking Computational XAI Evaluation to Human Understanding

2026-03-26 · Gregor Baer, Chao Zhang, Isel Grau, Pieter Van Gorp arxiv

Explainable AI (XAI) methods are commonly evaluated with functional metrics such as correctness, which computationally estimate how accurately an explanation reflects the model's reasoning. Higher correctness is assumed to produce better human understanding, but this link has not been tested experimentally with controlled levels. We conducted a user study (N=200) that manipulated explanation correctness at four levels (100%, 85%, 70%, 55%) in a time series classification task where participants could not rely on domain knowledge or visual intuition and instead predicted the AI's decisions based on explanations (forward simulation). Correctness affected understanding, but not at every level: performance dropped at 70% and 55% correctness relative to fully correct explanations, while further degradation below 70% produced no additional loss. Rather than shifting performance uniformly, lower correctness decreased the proportion of participants who learned the decision pattern. At the same time, even fully correct explanations did not guarantee understanding, as only a subset of participants achieved high accuracy. Exploratory analyses showed that self-reported ratings correlated with demonstrated performance only when explanations were fully correct and participants had learned the pattern. These findings show that not all differences in functional correctness translate to differences in human understanding, underscoring the need to validate functional metrics against human outcomes.

📄 PDF Abstract BibTeX arXiv:2603.25251

Code (0)

등록된 구현이 없습니다.

Tasks

Time Series Classification

Similar Papers 제목 키워드 기반

Advancing Post Hoc Case Based Explanation with Feature Highlighting

2023-11-06 · Eoin Kenny, Eoin Delaney, Mark Keane

Explainable AI (XAI) has been proposed as a valuable tool to assist in downstream tasks involving human and AI collaboration. Perhaps the most psychologically valid XAI techniques are case based approaches which display …

valid

Beyond saliency: enhancing explanation of speech emotion recognition with expert-referenced acoustic cues

2025-11-12 · Seham Nasr, Zhao Ren, David Johnson arxiv

Explainable AI (XAI) for Speech Emotion Recognition (SER) is critical for building transparent, trustworthy models. Current saliency-based methods, adapted from vision, highlight spectrogram regions but fail to show whet…

Speech Emotion Recognition

Generalizing Logic-based Explanations for Machine Learning Classifiers via Optimization

2026-03-02 · Francisco Mateus Rocha Filho, Ajalmar Rêgo da Rocha Neto, Thiago Alves Rocha arxiv

Machine learning models support decision-making, yet the reasons behind their predictions are opaque. Clear and reliable explanations help users make informed decisions and avoid blindly trusting model outputs. However, …

An Evaluation of the Human-Interpretability of Explanation

2019-01-31 · Isaac Lage, Emily Chen, Jeffrey He, Menaka Narayanan 외

Recent years have seen a boom in interest in machine learning systems that can provide a human-understandable rationale for their predictions or decisions. However, exactly what kinds of explanation are truly human-inter…

BIG-bench Machine Learning

Linking Uncertainty in Physicians' Narratives to Diagnostic Correctness

2012-07-01 · WS 2012 7 · Wilson McCoy, Cecilia Ovesdotter Alm, Cara Calvelli, Jeff B. Pelz 외
Decision MakingDiagnostic