paper-with-me

Papers

A Study of F0 Modification for X-Vector Based Speech Pseudonymization Across Gender

2021-01-21 · Pierre Champion, Denis Jouvet, Anthony Larcher

Speech pseudonymization aims at altering a speech signal to map the identifiable personal characteristics of a given speaker to another identity. In other words, it aims to hide the source speaker identity while preserving the intelligibility of the spoken content. This study takes place in the VoicePrivacy 2020 challenge framework, where the baseline system performs pseudonymization by modifying x-vector information to match a target speaker while keeping the fundamental frequency (F0) unchanged. We propose to alter other paralin-guistic features, here F0, and analyze the impact of this modification across gender. We found that the proposed F0 modification always improves pseudonymization We observed that both source and target speaker genders affect the performance gain when modifying the F0.

📄 PDF Abstract BibTeX arXiv:2101.08478

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Privacy- and Utility-Preserving NLP with Anonymized Data: A case study of Pseudonymization

2023-06-08 · Oleksandr Yermilov, Vipul Raheja, Artem Chernodub

This work investigates the effectiveness of different pseudonymization techniques, ranging from rule-based substitutions to using pre-trained Large Language Models (LLMs), on a variety of datasets and models used for two…

text-classificationText Classification

Grandma Karl is 27 years old -- research agenda for pseudonymization of research data

2023-08-30 · Elena Volodina, Simon Dobnik, Therese Lindström Tiedemann, Xuan-Son Vu

Accessibility of research data is critical for advances in many research fields, but textual data often cannot be shared due to the personal and sensitive information which it contains, e.g names or political opinions. G…

AnonShield: Scalable On-Premise Pseudonymization for CSIRT Vulnerability Data

2026-04-05 · Cristhian Kapelinski, Douglas Lautert, Beatriz Machado, Diego Kreutz 외 arxiv

We present AnonShield, a high-throughput, on-premise pseudonymization system that combines GPU-accelerated NER, streaming processing, caching, and schema-aware configuration. Evaluated on datasets up to 550 MB (70,951 re…

Towards Privacy by Design in Learner Corpora Research: A Case of On-the-fly Pseudonymization of Swedish Learner Essays

2020-12-01 · COLING 2020 8 · Elena Volodina, Yousuf Ali Mohammed, Sandra Derbring, Arild Matsson 외

This article reports on an ongoing project aiming at automatization of pseudonymization of learner essays. The process includes three steps: identification of personal information in an unstructured text, labeling for a …

Decomposing Memorization Reduction in Privacy-Preserving Fine-Tuning of SLMs for CSIRTs

2026-06-26 · Cristhian Kapelinski, Diego Kreutz arxiv

CSIRTs increasingly fine tune language models on vulnerability scan records, but these records expose internal network topology and create privacy risks under regulations such as GDPR and LGPD. We present the first empir…