Spoof Detection
2개 벤치마크 · 논문 7편 · 이 태스크의 논문 보기 →
Benchmarks
ASVspoof 2019 - LA
ASVspoof 2019 - PA
Most implemented
Papers
Supervised Post-training of Speech Foundation Models for Robust Adaptation in Speech Deepfake Detection
Large speech foundation models have shown strong potential for speech deepfake detection, but direct fine-tuning is limited by a mismatch between self-supervised pre-training objectives and spoof-specific artifacts. To a…
DeepFake DetectionData AugmentationSpoof DetectionIllumination-Aware Contactless Fingerprint Spoof Detection via Paired Flash-Non-Flash Imaging
Contactless fingerprint recognition enables hygienic and convenient biometric authentication but poses new challenges for spoof detection due to the absence of physical contact and traditional liveness cues. Most existin…
Spoof DetectionHow to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection
Since Text-to-Speech systems typically don't produce waveforms directly, recent spoof detection studies use resynthesized waveforms from vocoders and neural audio codecs to simulate an attacker. Unlike vocoders, which ar…
Audio Deepfake DetectionSpeech SynthesisSpoof DetectionConditional Synthetic Live and Spoof Fingerprint Generation
Large fingerprint datasets, while important for training and evaluation, are time-consuming and expensive to collect and require strict privacy measures. Researchers are exploring the use of synthetic fingerprint data to…
Spoof DetectionArFake: A Robust Framework for Multi-Dialect Arabic Speech Spoofing Detection Benchmark
With the rise of generative text-to-speech models, distinguishing between real and synthetic speech has become challenging, especially for Arabic that have received limited research attention. Most spoof detection effort…
Speech RecognitionSpoof DetectionHuLA: Prosody-Aware Anti-Spoofing with Multi-Task Learning for Expressive and Emotional Synthetic Speech
Current anti-spoofing systems remain vulnerable to expressive and emotional synthetic speech, since they rarely leverage prosody as a discriminative cue. Prosody is central to human expressiveness and emotion, and humans…
Self-Supervised LearningMulti-Task LearningSpoof Detection