paper-with-me

홈 › Papers

SpecRNet: Towards Faster and More Accessible Audio DeepFake Detection

2022-10-12 · Piotr Kawa, Marcin Plata, Piotr Syga

Audio DeepFakes are utterances generated with the use of deep neural networks. They are highly misleading and pose a threat due to use in fake news, impersonation, or extortion. In this work, we focus on increasing accessibility to the audio DeepFake detection methods by providing SpecRNet, a neural network architecture characterized by a quick inference time and low computational requirements. Our benchmark shows that SpecRNet, requiring up to about 40% less time to process an audio sample, provides performance comparable to LCNN architecture - one of the best audio DeepFake detection models. Such a method can not only be used by online multimedia services to verify a large bulk of content uploaded daily but also, thanks to its low requirements, by average citizens to evaluate materials on their devices. In addition, we provide benchmarks in three unique settings that confirm the correctness of our model. They reflect scenarios of low-resource datasets, detection on short utterances and limited attacks benchmark in which we take a closer look at the influence of particular attacks on given architectures.

📄 PDF Abstract BibTeX arXiv:2210.06105

Code (1)

piotrkawa/specrnet 공식 구현 pytorch

Tasks

Audio Deepfake DetectionDeepFake DetectionFace Swapping

Similar Papers 제목 키워드 기반

Improved DeepFake Detection Using Whisper Features

2023-06-02 · Piotr Kawa, Marcin Plata, Michał Czuba, Piotr Szymański 외

With a recent influx of voice generation methods, the threat introduced by audio DeepFake (DF) is ever-increasing. Several different detection methods have been presented as a countermeasure. Many methods are based on so…

Automatic Speech RecognitionDeepFake DetectionFace Swappingspeech-recognition+1

Multi-Speaker Conversational Audio Deepfake: Taxonomy, Dataset and Pilot Study

2026-01-30 · Alabi Ahmed, Vandana Janeja, Sanjay Purushotham arxiv

The rapid advances in text-to-speech (TTS) technologies have made audio deepfakes increasingly realistic and accessible, raising significant security and trust concerns. While existing research has largely focused on det…

DeepFake Detection

IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection

2025-06-23 · Abhay Kumar, Kunal Verma, Omkar More

Advancements in audio deepfake technology offers benefits like AI assistants, better accessibility for speech impairments, and enhanced entertainment. However, it also poses significant risks to security, privacy, and tr…

Audio Deepfake DetectionDeepFake DetectionFace Swapping

Interpretable All-Type Audio Deepfake Detection with Audio LLMs via Frequency-Time Reinforcement Learning

2026-01-06 · Yuankun Xie, Xiaoxuan Guo, Jiayi Zhou, Tao Wang 외 arxiv

Recent advances in audio large language models (ALLMs) have made high-quality synthetic audio widely accessible, increasing the risk of malicious audio deepfakes across speech, environmental sounds, singing voice, and mu…

Audio Deepfake DetectionReinforcement Learning

FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

2021-08-11 · Hasam Khalid, Shahroz Tariq, Minha Kim, Simon S. Woo

While the significant advancements have made in the generation of deepfakes using deep learning technologies, its misuse is a well-known issue now. Deepfakes can cause severe security and privacy issues as they can be us…

DeepFake DetectionFace Swapping