paper-with-me

Papers

The Impact of Silence on Speech Anti-Spoofing

2023-09-21 · Yuxiang Zhang, Zhuo Li, Jingze Lu, Hua Hua, Wenchao Wang, Pengyuan Zhang

The current speech anti-spoofing countermeasures (CMs) show excellent performance on specific datasets. However, removing the silence of test speech through Voice Activity Detection (VAD) can severely degrade performance. In this paper, the impact of silence on speech anti-spoofing is analyzed. First, the reasons for the impact are explored, including the proportion of silence duration and the content of silence. The proportion of silence duration in spoof speech generated by text-to-speech (TTS) algorithms is lower than that in bonafide speech. And the content of silence generated by different waveform generators varies compared to bonafide speech. Then the impact of silence on model prediction is explored. Even after retraining, the spoof speech generated by neural network based end-to-end TTS algorithms suffers a significant rise in error rates when the silence is removed. To demonstrate the reasons for the impact of silence on CMs, the attention distribution of a CM is visualized through class activation mapping (CAM). Furthermore, the implementation and analysis of the experiments masking silence or non-silence demonstrates the significance of the proportion of silence duration for detecting TTS and the importance of silence content for detecting voice conversion (VC). Based on the experimental results, improving the robustness of CMs against unknown spoofing attacks by masking silence is also proposed. Finally, the attacks on anti-spoofing CMs through concatenating silence, and the mitigation of VAD and silence attack through low-pass filtering are introduced.

📄 PDF Abstract BibTeX arXiv:2309.11827

Code (0)

등록된 구현이 없습니다.

Tasks

Action DetectionActivity Detectiontext-to-speechText to SpeechVoice Conversion

Similar Papers 제목 키워드 기반

Beyond Silence: Bias Analysis through Loss and Asymmetric Approach in Audio Anti-Spoofing

2024-06-25 · Hye-jin Shim, Md Sahidullah, Jee-weon Jung, Shinji Watanabe 외

Current trends in audio anti-spoofing detection research strive to improve models' ability to generalize across unseen attacks by learning to identify a variety of spoofing artifacts. This emphasis has primarily focused …

The MSXF TTS System for ICASSP 2022 ADD Challenge

2022-01-27 · Chunyong Yang, PengFei Liu, Yanli Chen, Hongbin Wang 외

This paper presents our MSXF TTS system for Task 3.1 of the Audio Deep Synthesis Detection (ADD) Challenge 2022. We use an end to end text to speech system, and add a constraint loss to the system when training stage. Th…

text-to-speechText to Speech

Thech. Report: Genuinization of Speech waveform PMF for speaker detection spoofing and countermeasures

2023-10-09 · Itshak Lapidot, Jean-Francois Bonastre

In the context of spoofing attacks in speaker recognition systems, we observed that the waveform probability mass function (PMF) of genuine speech differs significantly from the PMF of speech resulting from the attacks. …

Speaker Recognition

The Impact of Audio Watermarking on Audio Anti-Spoofing Countermeasures

2025-09-25 · Zhenshan Zhang, Xueping Zhang, Yechen Wang, Liwei Jin 외 arxiv

This paper presents the first study on the impact of audio watermarking on spoofing countermeasures. While anti-spoofing systems are essential for securing speech-based applications, the influence of widely used audio wa…

Synthetic Speech Detection Based on Temporal Consistency and Distribution of Speaker Features

2023-09-29 · Yuxiang Zhang, Zhuo Li, Jingze Lu, Wenchao Wang 외

Current synthetic speech detection (SSD) methods perform well on certain datasets but still face issues of robustness and interpretability. A possible reason is that these methods do not analyze the deficiencies of synth…

Synthetic Speech Detectiontext-to-speechText to Speech