Attack on practical speaker verification system using universal adversarial perturbations
In authentication scenarios, applications of practical speaker verification systems usually require a person to read a dynamic authentication text. Previous studies played an audio adversarial example as a digital signal to perform physical attacks, which would be easily rejected by audio replay detection modules. This work shows that by playing our crafted adversarial perturbation as a separate source when the adversary is speaking, the practical speaker verification system will misjudge the adversary as a target speaker. A two-step algorithm is proposed to optimize the universal adversarial perturbation to be text-independent and has little effect on the authentication text recognition. We also estimated room impulse response (RIR) in the algorithm which allowed the perturbation to be effective after being played over the air. In the physical experiment, we achieved targeted attacks with success rate of 100%, while the word error rate (WER) on speech recognition was only increased by 3.55%. And recorded audios could pass replay detection for the live person speaking.
Code (1)
Tasks
Real-World Adversarial AttackRoom Impulse Response (RIR)Speaker Verificationspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
MASTERKEY: Practical Backdoor Attack Against Speaker Verification Systems
Speaker Verification (SV) is widely deployed in mobile systems to authenticate legitimate users by using their voice traits. In this work, we propose a backdoor attack MASTERKEY, to compromise the SV models. Different fr…
Backdoor AttackSpeaker VerificationIntegrated Replay Spoofing-aware Text-independent Speaker Verification
A number of studies have successfully developed speaker verification or presentation attack detection systems. However, studies integrating the two tasks remain in the preliminary stages. In this paper, we propose two ap…
Multi-Task LearningSpeaker IdentificationSpeaker VerificationText-Independent Speaker VerificationUniversal speaker recognition encoders for different speech segments duration
Creating universal speaker encoders which are robust for different acoustic and speech duration conditions is a big challenge today. According to our observations systems trained on short speech segments are optimal for …
Speaker RecognitionSpeaker VerificationReal-time, Universal, and Robust Adversarial Attacks Against Speaker Recognition Systems
As the popularity of voice user interface (VUI) exploded in recent years, speaker recognition system has emerged as an important medium of identifying a speaker in many security-required applications and services. In thi…
Adversarial AttackRoom Impulse Response (RIR)Speaker RecognitionA Master Key Backdoor for Universal Impersonation Attack against DNN-based Face Verification
We introduce a new attack against face verification systems based on Deep Neural Networks (DNN). The attack relies on the introduction into the network of a hidden backdoor, whose activation at test time induces a verifi…
Backdoor AttackFace Verification