paper-with-me

홈 › Papers

Learning From Yourself: A Self-Distillation Method for Fake Speech Detection

2023-03-02 · Jun Xue, Cunhang Fan, Jiangyan Yi, Chenglong Wang, Zhengqi Wen, Dan Zhang, Zhao Lv

In this paper, we propose a novel self-distillation method for fake speech detection (FSD), which can significantly improve the performance of FSD without increasing the model complexity. For FSD, some fine-grained information is very important, such as spectrogram defects, mute segments, and so on, which are often perceived by shallow networks. However, shallow networks have much noise, which can not capture this very well. To address this problem, we propose using the deepest network instruct shallow network for enhancing shallow networks. Specifically, the networks of FSD are divided into several segments, the deepest network being used as the teacher model, and all shallow networks become multiple student models by adding classifiers. Meanwhile, the distillation path between the deepest network feature and shallow network features is used to reduce the feature difference. A series of experimental results on the ASVspoof 2019 LA and PA datasets show the effectiveness of the proposed method, with significant improvements compared to the baseline.

📄 PDF Abstract BibTeX arXiv:2303.01211

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Post-training for Deepfake Speech Detection

2025-06-26 · Wanying Ge, Xin Wang, Xuechen Liu, Junichi Yamagishi

We introduce a post-training approach that adapts self-supervised learning (SSL) models for deepfake speech detection by bridging the gap between general pre-training and domain-specific fine-tuning. We present AntiDeepf…

Face SwappingSelf-Supervised Learning

Fake-Mamba: Real-Time Speech Deepfake Detection Using Bidirectional Mamba as Self-Attention's Alternative

2025-08-12 · Xi Xuan, Zimo Zhu, Wenxin Zhang, Yi-Cheng Lin 외 arxiv

Advances in speech synthesis intensify security threats, motivating real-time deepfake detection research. We investigate whether bidirectional Mamba can serve as a competitive alternative to Self-Attention in detecting …

DeepFake DetectionSpeech Synthesis

Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform

2025-08-14 · Yuankun Xie, Ruibo Fu, Xiaopeng Wang, Zhiyong Wang 외 arxiv

The rapid advancement of speech generation technology has led to the widespread proliferation of deepfake speech across social media platforms. While deepfake audio countermeasures (CMs) achieve promising results on publ…

DeepFake DetectionData Augmentation

CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation

2025-05-21 · Yuxuan Du, Zhendong Wang, Yuhao Luo, Caiyong Piao 외

The rapid emergence of multimodal deepfakes (visual and auditory content are manipulated in concert) undermines the reliability of existing detectors that rely solely on modality-specific artifacts or cross-modal inconsi…

cross-modal alignmentDeepFake DetectionFace Swapping

Partially Fake Audio Detection by Self-attention-based Fake Span Discovery

2022-02-14 · Haibin Wu, Heng-Cheng Kuo, Naijun Zheng, Kuo-Hsuan Hung 외

The past few years have witnessed the significant advances of speech synthesis and voice conversion technologies. However, such technologies can undermine the robustness of broadly implemented biometric identification mo…

Open-Ended Question AnsweringQuestion AnsweringSpeech SynthesisVoice Conversion