paper-with-me

Papers

Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example

2024-10-20 · Suhita Ghosh, Melanie Jouaiti, Arnab Das, Yamini Sinha, Tim Polzehl, Ingo Siegert, Sebastian Stober

Speech anonymisation aims to protect speaker identity by changing personal identifiers in speech while retaining linguistic content. Current methods fail to retain prosody and unique speech patterns found in elderly and pathological speech domains, which is essential for remote health monitoring. To address this gap, we propose a voice conversion-based method (DDSP-QbE) using differentiable digital signal processing and query-by-example. The proposed method, trained with novel losses, aids in disentangling linguistic, prosodic, and domain representations, enabling the model to adapt to uncommon speech patterns. Objective and subjective evaluations show that DDSP-QbE significantly outperforms the voice conversion state-of-the-art concerning intelligibility, prosody, and domain preservation across diverse datasets, pathologies, and speakers while maintaining quality and speaker anonymity. Experts validate domain preservation by analysing twelve clinically pertinent domain attributes.

📄 PDF Abstract BibTeX arXiv:2410.15500

Code (1)

suhitaghosh10/ddsp-qbe 공식 구현 pytorch

Tasks

Voice Conversion

Similar Papers 제목 키워드 기반

Pathological voice adaptation with autoencoder-based voice conversion

2021-06-15 · Marc Illa, Bence Mark Halpern, Rob van Son, Laureano Moro-Velazquez 외

In this paper, we propose a new approach to pathological speech synthesis. Instead of using healthy speech as a source, we customise an existing pathological speech sample to a new speaker's voice characteristics. This a…

Speech SynthesisVoice Conversion

An Objective Evaluation Framework for Pathological Speech Synthesis

2021-07-01 · Bence Mark Halpern, Julian Fritsch, Enno Hermann, Rob van Son 외

The development of pathological speech systems is currently hindered by the lack of a standardised objective evaluation framework. In this work, (1) we utilise existing detection and analysis techniques to propose a gene…

Speech SynthesisVoice Conversion

Self-Supervised Speech Representations Preserve Speech Characteristics while Anonymizing Voices

2022-04-04 · Abner Hernandez, Paula Andrea Pérez-Toro, Juan Camilo Vásquez-Correa, Juan Rafael Orozco-Arroyave 외

Collecting speech data is an important step in training speech recognition systems and other speech-based machine learning models. However, the issue of privacy protection is an increasing concern that must be addressed.…

Speaker Verificationspeech-recognitionSpeech RecognitionVoice Conversion

VOTE400(Voide Of The Elderly 400 Hours): A Speech Dataset to Study Voice Interface for Elderly-Care

2021-01-20 · Minsu Jang, Sangwon Seo, Dohyung Kim, Jaeyeon Lee 외

This paper introduces a large-scale Korean speech dataset, called VOTE400, that can be used for analyzing and recognizing voices of the elderly people. The dataset includes about 300 hours of continuous dialog speech and…

speech-recognitionSpeech Recognition

Towards Identity Preserving Normal to Dysarthric Voice Conversion

2021-10-15 · Wen-Chin Huang, Bence Mark Halpern, Lester Phillip Violeta, Odette Scharenborg 외

We present a voice conversion framework that converts normal speech into dysarthric speech while preserving the speaker identity. Such a framework is essential for (1) clinical decision making processes and alleviation o…

Data AugmentationDecision Makingspeech-recognitionSpeech Recognition+1