paper-with-me

Papers

Improving speaker de-identification with functional data analysis of f0 trajectories

2022-03-31 · Lauri Tavi, Tomi Kinnunen, Rosa González Hautamäki

Due to a constantly increasing amount of speech data that is stored in different types of databases, voice privacy has become a major concern. To respond to such concern, speech researchers have developed various methods for speaker de-identification. The state-of-the-art solutions utilize deep learning solutions which can be effective but might be unavailable or impractical to apply for, for example, under-resourced languages. Formant modification is a simpler, yet effective method for speaker de-identification which requires no training data. Still, remaining intonational patterns in formant-anonymized speech may contain speaker-dependent cues. This study introduces a novel speaker de-identification method, which, in addition to simple formant shifts, manipulates f0 trajectories based on functional data analysis. The proposed speaker de-identification method will conceal plausibly identifying pitch characteristics in a phonetically controllable manner and improve formant-based speaker de-identification up to 25%.

📄 PDF Abstract BibTeX arXiv:2203.16738

Code (1)

laurivel/de_id-fda-f0-formants 공식 구현

Tasks

De-identification

Similar Papers 제목 키워드 기반

Evaluating Identity Leakage in Speaker De-Identification Systems

2025-08-19 · Seungmin Seo, Oleg Aulov, Afzal Godil, Kevin Mangold arxiv

Speaker de-identification aims to conceal a speaker's identity while preserving intelligibility of the underlying speech. We introduce a benchmark that quantifies residual identity leakage with three complementary error …

Speaker Identification From Youtube Obtained Data

2014-11-11 · Nitesh Kumar Chaudhary

An efficient, and intuitive algorithm is presented for the identification of speakers from a long dataset (like YouTube long discussion, Cocktail party recorded audio or video).The goal of automatic speaker identificatio…

parameter estimationQuantizationSpeaker IdentificationSpeaker Verification

Native Language Identification with User Generated Content

2018-10-01 · EMNLP 2018 10 · Gili Goldin, Ella Rabinovich, Shuly Wintner

We address the task of native language identification in the context of social media content, where authors are highly-fluent, advanced nonnative speakers (of English). Using both linguistically-motivated features and th…

Language IdentificationNative Language Identification

Modular Deep Learning Framework for Assistive Perception: Gaze, Affect, and Speaker Identification

2025-11-25 · Akshit Pramod Anchan, Jewelith Thomas, Sritama Roy arxiv

Developing comprehensive assistive technologies requires the seamless integration of visual and auditory perception. This research evaluates the feasibility of a modular architecture inspired by core functionalities of p…

Facial Expression RecognitionSpeaker Identification

CASA-Based Speaker Identification Using Cascaded GMM-CNN Classifier in Noisy and Emotional Talking Conditions

2021-02-11 · Ali Bou Nassif, Ismail Shahin, Shibani Hamsa, Nawel Nemmour 외

This work aims at intensifying text-independent speaker identification performance in real application situations such as noisy and emotional talking conditions. This is achieved by incorporating two different modules: a…

Emotion RecognitionSpeaker Identification