paper-with-me

Papers

Anonymizing Speech with Generative Adversarial Networks to Preserve Speaker Privacy

2022-10-13 · Sarina Meyer, Pascal Tilli, Pavel Denisov, Florian Lux, Julia Koch, Ngoc Thang Vu

In order to protect the privacy of speech data, speaker anonymization aims for hiding the identity of a speaker by changing the voice in speech recordings. This typically comes with a privacy-utility trade-off between protection of individuals and usability of the data for downstream applications. One of the challenges in this context is to create non-existent voices that sound as natural as possible. In this work, we propose to tackle this issue by generating speaker embeddings using a generative adversarial network with Wasserstein distance as cost function. By incorporating these artificial embeddings into a speech-to-text-to-speech pipeline, we outperform previous approaches in terms of privacy and utility. According to standard objective metrics and human evaluation, our approach generates intelligible and content-preserving yet privacy-protecting versions of the original recordings.

📄 PDF Abstract BibTeX arXiv:2210.07002

Code (1)

digitalphonetics/speaker-anonymization 공식 구현 pytorch

Tasks

Generative Adversarial NetworkSpeaker anonymizationSpeech-to-Texttext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Self-Supervised Speech Representations Preserve Speech Characteristics while Anonymizing Voices

2022-04-04 · Abner Hernandez, Paula Andrea Pérez-Toro, Juan Camilo Vásquez-Correa, Juan Rafael Orozco-Arroyave 외

Collecting speech data is an important step in training speech recognition systems and other speech-based machine learning models. However, the issue of privacy protection is an increasing concern that must be addressed.…

Speaker Verificationspeech-recognitionSpeech RecognitionVoice Conversion

Asynchronous Voice Anonymization Using Adversarial Perturbation On Speaker Embedding

2024-06-12 · Rui Wang, Liping Chen, Kong Aik Lee, Zhen-Hua Ling

Voice anonymization has been developed as a technique for preserving privacy by replacing the speaker's voice in a speech signal with that of a pseudo-speaker, thereby obscuring the original voice attributes from machine…

Disentanglement

Anonymizing Speech: Evaluating and Designing Speaker Anonymization Techniques

2023-08-05 · Pierre Champion

The growing use of voice user interfaces has led to a surge in the collection and storage of speech data. While data collection allows for the development of efficient tools powering most speech services, it also poses s…

QuantizationSpeaker anonymizationVoice CloningVoice Conversion

Speaker Identity Preservation in Dysarthric Speech Reconstruction by Adversarial Speaker Adaptation

2022-02-18 · Disong Wang, Songxiang Liu, Xixin Wu, Hui Lu 외

Dysarthric speech reconstruction (DSR), which aims to improve the quality of dysarthric speech, remains a challenge, not only because we need to restore the speech to be normal, but also must preserve the speaker's ident…

Multi-Task LearningSpeaker Verification

Privacy versus Emotion Preservation Trade-offs in Emotion-Preserving Speaker Anonymization

2024-09-05 · Zexin Cai, Henry Li Xinyuan, Ashi Garg, Leibny Paola García-Perera 외

Advances in speech technology now allow unprecedented access to personally identifiable information through speech. To protect such information, the differential privacy field has explored ways to anonymize speech while …

Speaker anonymizationSpeaker Verification