paper-with-me

Papers

Exploiting Context-dependent Duration Features for Voice Anonymization Attack Systems

2025-07-21 · Natalia Tomashenko, Emmanuel Vincent, Marc Tommasi arxiv

The temporal dynamics of speech, encompassing variations in rhythm, intonation, and speaking rate, contain important and unique information about speaker identity. This paper proposes a new method for representing speaker characteristics by extracting context-dependent duration embeddings from speech temporal dynamics. We develop novel attack models using these representations and analyze the potential vulnerabilities in speaker verification and voice anonymization systems.The experimental results show that the developed attack models provide a significant improvement in speaker verification performance for both original and anonymized data in comparison with simpler representations of speech temporal dynamics reported in the literature.

📄 PDF Abstract BibTeX arXiv:2507.15214

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Similar Papers 제목 키워드 기반

Respiratory Distress Detection from Telephone Speech using Acoustic and Prosodic Features

2020-11-15 · Meemnur Rashid, Kaisar Ahmed Alman, Khaled Hasan, John H. L. Hansen 외

With the widespread use of telemedicine services, automatic assessment of health conditions via telephone speech can significantly impact public health. This work summarizes our preliminary findings on automatic detectio…

Singing-Tacotron: Global duration control attention and dynamic filter for End-to-end singing voice synthesis

2022-02-16 · Tao Wang, Ruibo Fu, Jiangyan Yi, JianHua Tao 외

End-to-end singing voice synthesis (SVS) is attractive due to the avoidance of pre-aligned data. However, the auto learned alignment of singing voice with lyrics is difficult to match the duration information in musical …

Singing Voice Synthesis

Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization

2024-12-22 · Natalia Tomashenko, Emmanuel Vincent, Marc Tommasi

In this paper, we investigate the impact of speech temporal dynamics in application to automatic speaker verification and speaker voice anonymization tasks. We propose several metrics to perform automatic speaker verific…

Speaker Verification

PromptVC: Flexible Stylistic Voice Conversion in Latent Space Driven by Natural Language Prompts

2023-09-17 · Jixun Yao, Yuguang Yang, Yi Lei, Ziqian Ning 외

Style voice conversion aims to transform the style of source speech to a desired style according to real-world application demands. However, the current style voice conversion approach relies on pre-defined labels or ref…

Voice Conversion

VoiceExtender: Short-utterance Text-independent Speaker Verification with Guided Diffusion Model

2023-10-07 · Yayun He, Zuheng Kang, Jianzong Wang, Junqing Peng 외

Speaker verification (SV) performance deteriorates as utterances become shorter. To this end, we propose a new architecture called VoiceExtender which provides a promising solution for improving SV performance when handl…

Speaker VerificationText-Independent Speaker Verification