paper-with-me

홈 › Papers

Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning

2026-03-11 · Artem Dvirniak, Evgeny Kushnir, Dmitrii Tarasov, Artem Iudin, Oleg Kiriukhin, Mikhail Pautov, Dmitrii Korzh, Oleg Y. Rogov arxiv

The modern generative audio models can be used by an adversary in an unlawful manner, specifically, to impersonate other people to gain access to private information. To mitigate this issue, speech deepfake detection (SDD) methods started to evolve. Unfortunately, current SDD methods generally suffer from the lack of generalization to new audio domains and generators. More than that, they lack interpretability, especially human-like reasoning that would naturally explain the attribution of a given audio to the bona fide or spoof class and provide human-perceptible cues. In this paper, we propose HIR-SDD, a novel SDD framework that combines the strengths of Large Audio Language Models (LALMs) with the chain-of-thought reasoning derived from the novel proposed human-annotated dataset. Experimental evaluation demonstrates both the effectiveness of the proposed method and its ability to provide reasonable justifications for predictions.

📄 PDF Abstract BibTeX arXiv:2603.10725

Code (0)

등록된 구현이 없습니다.

Tasks

DeepFake Detection

Similar Papers 제목 키워드 기반

Sparse deepfake detection promotes better disentanglement

2025-10-07 · Antoine Teissier, Marie Tahon, Nicolas Dugué, Aghilas Sini arxiv

Due to the rapid progress of speech synthesis, deepfake detection has become a major concern in the speech processing community. Because it is a critical task, systems must not only be efficient and robust, but also prov…

DeepFake DetectionSpeech Synthesis

REIMU: Efficient Heterogeneous Hierarchical Reasoning for SSL-Based Speech Deepfake Detection

2026-08-01 · Kwok-Ho Ng, Tingting Song, Bingwen Feng, Peiya Li arxiv

The increasing realism of speech generated by text-to-speech and voice conversion systems poses growing challenges to media integrity and voice authentication. Self-supervised learning (SSL) has substantially advanced sp…

Self-Supervised LearningDeepFake DetectionVoice Conversion

Toward Transdisciplinary Approaches to Audio Deepfake Discernment

2024-11-08 · Vandana P. Janeja, Christine Mallinson

This perspective calls for scholars across disciplines to address the challenge of audio deepfake detection and discernment through an interdisciplinary lens across Artificial Intelligence methods and linguistics. With a…

Audio Deepfake DetectionDeepFake DetectionFace Swapping

Quantum Vision Theory Applied to Audio Classification for Deepfake Speech Detection

2026-04-09 · Khalid Zaman, Melike Sah, Anuwat Chaiwongyenc, Cem Direkoglu arxiv

We propose Quantum Vision (QV) theory as a new perspective for deep learning-based audio classification, applied to deepfake speech detection. Inspired by particle-wave duality in quantum physics, QV theory is based on t…

Audio Deepfake DetectionAudio ClassificationImage Classification

AUDETER: A Large-scale Dataset for Deepfake Audio Detection in Open Worlds

2025-09-04 · Qizhou Wang, Hanxun Huang, Guansong Pang, Sarah Erfani 외 arxiv

Speech synthesis systems can now produce highly realistic vocalisations that pose significant authenticity challenges. Despite substantial progress in deepfake detection models, their real-world effectiveness is often un…

DeepFake DetectionSpeech Synthesis