paper-with-me

Papers

Beyond Silence: Bias Analysis through Loss and Asymmetric Approach in Audio Anti-Spoofing

2024-06-25 · Hye-jin Shim, Md Sahidullah, Jee-weon Jung, Shinji Watanabe, Tomi Kinnunen

Current trends in audio anti-spoofing detection research strive to improve models' ability to generalize across unseen attacks by learning to identify a variety of spoofing artifacts. This emphasis has primarily focused on the spoof class. Recently, several studies have noted that the distribution of silence differs between the two classes, which can serve as a shortcut. In this paper, we extend class-wise interpretations beyond silence. We employ loss analysis and asymmetric methodologies to move away from traditional attack-focused and result-oriented evaluations towards a deeper examination of model behaviors. Our investigations highlight the significant differences in training dynamics between the two classes, emphasizing the need for future research to focus on robust modeling of the bonafide class.

📄 PDF Abstract BibTeX arXiv:2406.17246

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Silenced Biases: The Dark Side LLMs Learned to Refuse

2025-11-05 · Rom Himelstein, Amit LeVi, Brit Youngmann, Yaniv Nemcovsky 외 arxiv

Safety-aligned large language models (LLMs) are becoming increasingly widespread, especially in sensitive applications where fairness is essential and biased outputs can cause significant harm. However, evaluating the fa…

The Impact of Silence on Speech Anti-Spoofing

2023-09-21 · Yuxiang Zhang, Zhuo Li, Jingze Lu, Hua Hua 외

The current speech anti-spoofing countermeasures (CMs) show excellent performance on specific datasets. However, removing the silence of test speech through Voice Activity Detection (VAD) can severely degrade performance…

Action DetectionActivity Detectiontext-to-speechText to Speech+1

Phonetic Error Analysis of Raw Waveform Acoustic Models with Parametric and Non-Parametric CNNs

2024-06-02 · Erfan Loweimi, Andrea Carmantini, Peter Bell, Steve Renals 외

In this paper, we analyse the error patterns of the raw waveform acoustic models in TIMIT's phone recognition task. Our analysis goes beyond the conventional phone error rate (PER) metric. We categorise the phones into t…

Transfer Learning

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

2024-11-29 · CVPR 2025 1 · Stefan Smeu, Dragos-Alexandru Boldisor, Dan Oneata, Elisabeta Oneata

Good datasets are essential for developing and benchmarking any machine learning system. Their importance is even more extreme for safety critical applications such as deepfake detection - the focus of this paper. Here w…

BenchmarkingDeepFake DetectionFace Swapping

DeepSilencer: A Novel Deep Learning Model for Predicting siRNA Knockdown Efficiency

2025-03-06 · Wangdan Liao, Weidong Wang

Background: Small interfering RNA (siRNA) is a promising therapeutic agent due to its ability to silence disease-related genes via RNA interference. While traditional machine learning and early deep learning methods have…

Deep Learning