paper-with-me

Papers

Hear No Evil: Towards Adversarial Robustness of Automatic Speech Recognition via Multi-Task Learning

2022-04-05 · Nilaksh Das, Duen Horng Chau

As automatic speech recognition (ASR) systems are now being widely deployed in the wild, the increasing threat of adversarial attacks raises serious questions about the security and reliability of using such systems. On the other hand, multi-task learning (MTL) has shown success in training models that can resist adversarial attacks in the computer vision domain. In this work, we investigate the impact of performing such multi-task learning on the adversarial robustness of ASR models in the speech domain. We conduct extensive MTL experimentation by combining semantically diverse tasks such as accent classification and ASR, and evaluate a wide range of adversarial settings. Our thorough analysis reveals that performing MTL with semantically diverse tasks consistently makes it harder for an adversarial attack to succeed. We also discuss in detail the serious pitfalls and their related remedies that have a significant impact on the robustness of MTL models. Our proposed MTL approach shows considerable absolute improvements in adversarially targeted WER ranging from 17.25 up to 59.90 compared to single-task learning baselines (attention decoder and CTC respectively). Ours is the first in-depth study that uncovers adversarial robustness gains from multi-task learning for ASR.

📄 PDF Abstract BibTeX arXiv:2204.02381

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial AttackAdversarial RobustnessAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderMulti-Task Learningspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

On the robustness of non-intrusive speech quality model by adversarial examples

2022-11-11 · Hsin-Yi Lin, Huan-Hsin Tseng, Yu Tsao

It has been shown recently that deep learning based models are effective on speech quality prediction and could outperform traditional metrics in various perspectives. Although network models have potential to be a surro…

Prediction

Did you hear that? Adversarial Examples Against Automatic Speech Recognition

2018-01-02 · Moustafa Alzantot, Bharathan Balaji, Mani Srivastava

Speech is a common and effective way of communication between humans, and modern consumer devices such as smartphones and home hubs are equipped with deep learning based accurate automatic speech recognition to enable na…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)object-detectionObject Detection+2

AeGAN: Time-Frequency Speech Denoising via Generative Adversarial Networks

2019-10-21 · Sherif Abdulatif, Karim Armanious, Karim Guirguis, Jayasankar T. Sajeev 외

Automatic speech recognition (ASR) systems are of vital importance nowadays in commonplace tasks such as speech-to-text processing and language translation. This created the need for an ASR system that can operate in rea…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingGenerative Adversarial Network+6

Reassessing Noise Augmentation Methods in the Context of Adversarial Speech

2024-09-03 · Karla Pizzi, Matías Pizarro, Asja Fischer

In this study, we investigate if noise-augmented training can concurrently improve adversarial robustness in automatic speech recognition (ASR) systems. We conduct a comparative analysis of the adversarial robustness of …

Adversarial RobustnessAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentation+2

Hear "No Evil", See "Kenansville": Efficient and Transferable Black-Box Attacks on Speech Recognition and Voice Identification Systems

2019-10-11 · Hadi Abdullah, Muhammad Sajidur Rahman, Washington Garcia, Logan Blue 외

Automatic speech recognition and voice identification systems are being deployed in a wide array of applications, from providing control mechanisms to devices lacking traditional interfaces, to the automatic transcriptio…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition