paper-with-me

Papers

Speaker Anonymization with Phonetic Intermediate Representations

2022-07-11 · Sarina Meyer, Florian Lux, Pavel Denisov, Julia Koch, Pascal Tilli, Ngoc Thang Vu

In this work, we propose a speaker anonymization pipeline that leverages high quality automatic speech recognition and synthesis systems to generate speech conditioned on phonetic transcriptions and anonymized speaker embeddings. Using phones as the intermediate representation ensures near complete elimination of speaker identity information from the input while preserving the original phonetic content as much as possible. Our experimental results on LibriSpeech and VCTK corpora reveal two key findings: 1) although automatic speech recognition produces imperfect transcriptions, our neural speech synthesis system can handle such errors, making our system feasible and robust, and 2) combining speaker embeddings from different resources is beneficial and their appropriate normalization is crucial. Overall, our final best system outperforms significantly the baselines provided in the Voice Privacy Challenge 2020 in terms of privacy robustness against a lazy-informed attacker while maintaining high intelligibility and naturalness of the anonymized speech.

📄 PDF Abstract BibTeX arXiv:2207.04834

Code (1)

digitalphonetics/speaker-anonymization 공식 구현 pytorch

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speaker anonymizationspeech-recognitionSpeech RecognitionSpeech Synthesis

Similar Papers 제목 키워드 기반

Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization

2024-12-22 · Natalia Tomashenko, Emmanuel Vincent, Marc Tommasi

In this paper, we investigate the impact of speech temporal dynamics in application to automatic speaker verification and speaker voice anonymization tasks. We propose several metrics to perform automatic speaker verific…

Speaker Verification

NWPU-ASLP System for the VoicePrivacy 2022 Challenge

2022-09-24 · Jixun Yao, Qing Wang, Li Zhang, Pengcheng Guo 외

This paper presents the NWPU-ASLP speaker anonymization system for VoicePrivacy 2022 Challenge. Our submission does not involve additional Automatic Speaker Verification (ASV) model or x-vector pool. Our system consists …

Speaker anonymizationSpeaker Verification

Why disentanglement-based speaker anonymization systems fail at preserving emotions?

2025-01-22 · Ünal Ege Gaznepoglu, Nils Peters

Disentanglement-based speaker anonymization involves decomposing speech into a semantically meaningful representation, altering the speaker embedding, and resynthesizing a waveform using a neural vocoder. State-of-the-ar…

DisentanglementEmotion RecognitionSpeaker anonymization

Reprogramming Self-supervised Learning-based Speech Representations for Speaker Anonymization

2023-11-17 · Xiaojiao Chen, Sheng Li, Jiyi Li, Hao Huang 외

Current speaker anonymization methods, especially with self-supervised learning (SSL) models, require massive computational resources when hiding speaker identity. This paper proposes an effective and parameter-efficient…

Self-Supervised LearningSpeaker anonymization

Automatic Voice Identification after Speech Resynthesis using PPG

2024-08-05 · Thibault Gaudier, Marie Tahon, Anthony Larcher, Yannick Estève

Speech resynthesis is a generic task for which we want to synthesize audio with another audio as input, which finds applications for media monitors and journalists.Among different tasks addressed by speech resynthesis, v…

ResynthesisSpeaker VerificationVoice Conversion