paper-with-me

Papers

Introducing Model Inversion Attacks on Automatic Speaker Recognition

2023-01-09 · Karla Pizzi, Franziska Boenisch, Ugur Sahin, Konstantin Böttinger

Model inversion (MI) attacks allow to reconstruct average per-class representations of a machine learning (ML) model's training data. It has been shown that in scenarios where each class corresponds to a different individual, such as face classifiers, this represents a severe privacy risk. In this work, we explore a new application for MI: the extraction of speakers' voices from a speaker recognition system. We present an approach to (1) reconstruct audio samples from a trained ML model and (2) extract intermediate voice feature representations which provide valuable insights into the speakers' biometrics. Therefore, we propose an extension of MI attacks which we call sliding model inversion. Our sliding MI extends standard MI by iteratively inverting overlapping chunks of the audio samples and thereby leveraging the sequential properties of audio data for enhanced inversion performance. We show that one can use the inverted audio data to generate spoofed audio samples to impersonate a speaker, and execute voice-protected commands for highly secured systems on their behalf. To the best of our knowledge, our work is the first one extending MI attacks to audio data, and our results highlight the security risks resulting from the extraction of the biometric data in that setup.

📄 PDF Abstract BibTeX arXiv:2301.03206

Code (0)

등록된 구현이 없습니다.

Tasks

modelSpeaker Recognition

Similar Papers 제목 키워드 기반

Scores Know Bobs Voice: Speaker Impersonation Attack

2026-03-03 · Chanwoo Hwang, Sunpill Kim, Yong Kiam Tan, Tianchi Liu 외 arxiv

Advances in deep learning have enabled the widespread deployment of speaker recognition systems (SRSs), yet they remain vulnerable to score-based impersonation attacks. Existing attacks that operate directly on raw wavef…

Speaker Recognition

Impact of Phonetics on Speaker Identity in Adversarial Voice Attack

2025-09-18 · Daniyal Kabir Dar, Qiben Yan, Li Xiao, Arun Ross arxiv

Adversarial perturbations in speech pose a serious threat to automatic speech recognition (ASR) and speaker verification by introducing subtle waveform modifications that remain imperceptible to humans but can significan…

Speaker VerificationSpeaker RecognitionSpeech Recognition

SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems

2020-07-13 · Hadi Abdullah, Kevin Warren, Vincent Bindschaedler, Nicolas Papernot 외

Speech and speaker recognition systems are employed in a variety of applications, from personal assistants to telephony surveillance and biometric authentication. The wide deployment of these systems has been made possib…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speaker IdentificationSpeaker Recognition+2

Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers

2020-06-19 · Naoyuki Kanda, Yashesh Gaur, Xiaofei Wang, Zhong Meng 외

We propose an end-to-end speaker-attributed automatic speech recognition model that unifies speaker counting, speech recognition, and speaker identification on monaural overlapped speech. Our model is built on serialized…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderSpeaker Identification+2

Exploiting Cross Domain Acoustic-to-articulatory Inverted Features For Disordered Speech Recognition

2022-03-19 · Shujie Hu, Shansong Liu, Xurong Xie, Mengzhe Geng 외

Articulatory features are inherently invariant to acoustic signal distortion and have been successfully incorporated into automatic speech recognition (ASR) systems for normal speech. Their practical application to disor…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+1