paper-with-me

홈 › Papers

An Information-Theoretic Framework for Comparing Voice and Text Explainability

2026-02-06 · Mona Rajhans, Vishal Khawarey arxiv

Explainable Artificial Intelligence (XAI) aims to make machine learning models transparent and trustworthy, yet most current approaches communicate explanations visually or through text. This paper introduces an information theoretic framework for analyzing how explanation modality specifically, voice versus text affects user comprehension and trust calibration in AI systems. The proposed model treats explanation delivery as a communication channel between model and user, characterized by metrics for information retention, comprehension efficiency (CE), and trust calibration error (T CE). A simulation framework implemented in Python was developed to evaluate these metrics using synthetic SHAP based feature attributions across multiple modality style configurations (brief, detailed, and analogy based). Results demonstrate that text explanations achieve higher comprehension efficiency, while voice explanations yield improved trust calibration, with analogy based delivery achieving the best overall trade off. This framework provides a reproducible foundation for designing and benchmarking multimodal explainability systems and can be extended to empirical studies using real SHAP or LIME outputs on open datasets such as the UCI Credit Approval or Kaggle Financial Transactions datasets.

📄 PDF Abstract BibTeX arXiv:2602.07179

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EEG-Derived Voice Signature for Attended Speaker Detection

2023-08-28 · Hongxu Zhu, Siqi Cai, Yidi Jiang, Qiquan Zhang 외

\textit{Objective:} Conventional EEG-based auditory attention detection (AAD) is achieved by comparing the time-varying speech stimuli and the elicited EEG signals. However, in order to obtain reliable correlation values…

EEG

Using Deepfake Technologies for Word Emphasis Detection

2023-05-12 · Eran Kaufman, Lee-Ad Gottlieb

In this work, we consider the task of automated emphasis detection for spoken language. This problem is challenging in that emphasis is affected by the particularities of speech of the subject, for example the subject ac…

Face Swapping

The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods

2018-04-12 · Jaime Lorenzo-Trueba, Junichi Yamagishi, Tomoki Toda, Daisuke Saito 외

We present the Voice Conversion Challenge 2018, designed as a follow up to the 2016 edition with the aim of providing a common framework for evaluating and comparing different state-of-the-art voice conversion (VC) syste…

Voice Conversion

Better Adherence, Richer Context: A Field Evaluation of LLM-Powered Conversational Voice Diaries for Sleep

2026-06-17 · Amama Mahmood, Bokyung Kim, Honghao Zhao, Molly E. Atwood 외 arxiv

Sleep diaries are central to behavioral sleep medicine and cognitive behavioral therapy for insomnia, yet daily completion is difficult to sustain, and static forms often provide limited context for interpreting night-to…

PPG-based singing voice conversion with adversarial representation learning

2020-10-28 · Zhonghao Li, Benlai Tang, Xiang Yin, Yuan Wan 외

Singing voice conversion (SVC) aims to convert the voice of one singer to that of other singers while keeping the singing content and melody. On top of recent voice conversion works, we propose a novel model to steadily …

Representation LearningVoice ConversionVoice Similarity