paper-with-me

홈 › Papers

Evolutionary Multi-Objective Fusion of Deepfake Speech Detectors

2026-04-01 · Vojtěch Staněk, Martin Perešíni, Lukáš Sekanina, Anton Firc, Kamil Malinka arxiv

While deepfake speech detectors built on large self-supervised learning (SSL) models achieve high accuracy, employing standard ensemble fusion to further enhance robustness often results in oversized systems with diminishing returns. To address this, we propose an evolutionary multi-objective score fusion framework that jointly minimizes detection error and system complexity. We explore two encodings optimized by NSGA-II: binary-coded detector selection for score averaging and a real-valued scheme that optimizes detector weights for a weighted sum. Experiments on the ASVspoof 5 dataset with 36 SSL-based detectors show that the obtained Pareto fronts outperform simple averaging and logistic regression baselines. The real-valued variant achieves 2.37% EER (0.0684 minDCF) and identifies configurations that match state-of-the-art performance while significantly reducing system complexity, requiring only half the parameters. Our method also provides a diverse set of trade-off solutions, enabling deployment choices that balance accuracy and computational cost.

📄 PDF Abstract BibTeX arXiv:2604.01330

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised Learning

Similar Papers 제목 키워드 기반

Diffuse or Confuse: A Diffusion Deepfake Speech Dataset

2024-10-09 · Anton Firc, Kamil Malinka, Petr Hanáček

Advancements in artificial intelligence and machine learning have significantly improved synthetic speech generation. This paper explores diffusion models, a novel method for creating realistic synthetic speech. We creat…

DeepFake DetectionFace Swapping

ExpSpeech-Net: Multimodal Fusion of Expression and Speech for Deepfake Detection

2026-06-04 · Ruchika Sharma, Rudresh Dwivedi arxiv

Deepfake videos are increasingly challenging the credibility of online content. Many existing detection methodology relies on complex, resource-intensive models, which limit their practical use. The study introduces the …

DeepFake Detection

A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection

2026-03-02 · Hashim Ali, Nithin Sai Adupa, Surya Subramani, Hafiz Malik arxiv

Self-supervised learning (SSL) has transformed speech processing, with benchmarks such as SUPERB establishing fair comparisons across diverse downstream tasks. Despite it's security-critical importance, Audio deepfake de…

Self-Supervised LearningAudio Deepfake Detection

Deepfake Audio Detection Using Self-supervised Fusion Representations

2026-05-05 · Khalid Zaman, Qixuan Huang, Muhammad Uzair, Masashi Unoki arxiv

This paper describes a submission to the Environment-Aware Speech and Sound Deepfake Detection Challenge (ESDD2) 2026, which addresses component-level deepfake detection using the CompSpoofV2 dataset, where speech and en…

DeepFake Detection

WavLM model ensemble for audio deepfake detection

2024-08-14 · David Combei, Adriana Stan, Dan Oneata, Horia Cucu

Audio deepfake detection has become a pivotal task over the last couple of years, as many recent speech synthesis and voice cloning systems generate highly realistic speech samples, thus enabling their use in malicious a…

Audio Deepfake DetectionData AugmentationDeepFake DetectionFace Swapping+3