paper-with-me

홈 › Papers

XAI-Grounded Explanation Generation for Speech Deepfake Detection with Training-Free Multimodal Large Language Models

2026-06-15 · Yupei Li, Qiyang Sun, Xiaoliang Wu, Chenxi Wang, Berrak Sisman, Björn W. Schuller arxiv

Speech deepfake detection (SDD) systems require trustworthy explanations for reliable decision-making. Existing explanation ways mainly fall into two categories. Traditional explainable AI (XAI), such as gradient-based attribution, produces low-level attribution signals tightly coupled with model decisions, and harder to be understood by human than natural language explanations. Meanwhile, large language model (LLM)-based explanation generation often produces generic and ungrounded descriptions due to the lack of heuristic evidence and task-specific supervision, stemming from limited grounded explanation datasets for SDD. We therefore propose a training-free explanation framework that integrates XAI evidence with multimodal LLMs to generate grounded and specific explanations. Using the PartialSpoof dataset, we construct a grounded explanation dataset and show that methods with XAI increase inside accuracy by over 45\%, verified through human evaluation and faithfulness checks.

📄 PDF Abstract BibTeX arXiv:2606.16137

Code (0)

등록된 구현이 없습니다.

Tasks

Explanation GenerationDeepFake Detection

Similar Papers 제목 키워드 기반

Explainable Deepfake Detection Challenge

2026-07-23 · Abhijeet Narang, Kartik Kuckreja, Shreya Ghosh, Muhammad Haris Khan 외 arxiv

Deepfake detection is moving beyond binary classification decisions toward systems that can also explain the visual evidence supporting those decisions. This transition is important for real-world verification settings, …

Explanation GenerationBinary ClassificationImage ClassificationSemantic Similarity

XPlainVerse: A Million-Scale Benchmark for Explainable Deepfake Detection

2026-07-03 · Abhijeet Narang, Kartik Kuckreja, Shreya Ghosh, Muhammad Haris Khan 외 arxiv

As deepfake detection models increasingly produce natural language explanations, their reasoning often remains weakly grounded in visual artifacts, limiting reliability and user trust. Existing benchmarks mainly evaluate…

DeepFake DetectionImage Editing

Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs

2025-09-26 · Xingyu Fu, Siyi Liu, Yinuo Xu, Pan Lu 외 arxiv

Can humans identify AI-generated (fake) videos and provide grounded reasons? While video generation models have advanced rapidly, a critical dimension -- whether humans can detect deepfake traces within a generated video…

Video Generation

Explainable Deepfake Detection with Feature-robust Augmentation and Evidence-grounded Explanation Optimization

2026-08-21 · Zhu Xu, Jiaqi Tang, Pokai Chen, Yuxin Peng 외 arxiv

Explainable deepfake detection extends binary classification by requiring models to not only predict authenticity but also provide interpretable justifications. This expanded scope is critical in practice, where users li…

Binary ClassificationContrastive LearningDeepFake Detection

Sparse deepfake detection promotes better disentanglement

2025-10-07 · Antoine Teissier, Marie Tahon, Nicolas Dugué, Aghilas Sini arxiv

Due to the rapid progress of speech synthesis, deepfake detection has become a major concern in the speech processing community. Because it is a critical task, systems must not only be efficient and robust, but also prov…

DeepFake DetectionSpeech Synthesis