paper-with-me

홈 › Papers

Addressing speaker gender bias in large scale speech translation systems

2025-01-10 · Shubham Bansal, Vikas Joshi, Harveen Chadha, Rupeshkumar Mehta, Jinyu Li

This study addresses the issue of speaker gender bias in Speech Translation (ST) systems, which can lead to offensive and inaccurate translations. The masculine bias often found in large-scale ST systems is typically perpetuated through training data derived from Machine Translation (MT) systems. Our approach involves two key steps. First, we employ Large Language Models (LLMs) to rectify translations based on the speaker's gender in a cost-effective manner. Second, we fine-tune the ST model with the corrected data, enabling the model to generate gender-specific translations directly from audio cues, without the need for explicit gender input. Additionally, we propose a three-mode fine-tuned model for scenarios where the speaker's gender is either predefined or should not be inferred from speech cues. We demonstrate a 70% improvement in translations for female speakers compared to our baseline and other large-scale ST systems, such as Seamless M4T and Canary, on the MuST-SHE test set.

📄 PDF Abstract BibTeX arXiv:2501.05989

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

No Pitch Left Behind: Addressing Gender Unbalance in Automatic Speech Recognition through Pitch Manipulation

2023-10-10 · Dennis Fucci, Marco Gaido, Matteo Negri, Mauro Cettolo 외

Automatic speech recognition (ASR) systems are known to be sensitive to the sociolinguistic variability of speech data, in which gender plays a crucial role. This can result in disparities in recognition accuracy between…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+1

Who Gets the Mic? Investigating Gender Bias in the Speaker Assignment of a Speech-LLM

2025-08-19 · Dariia Puhach, Amir H. Payberah, Éva Székely arxiv

Similar to text-based Large Language Models (LLMs), Speech-LLMs exhibit emergent abilities and context awareness. However, whether these similarities extend to gender bias remains an open question. This study proposes a …

Multi-Dimensional Gender Bias Classification

2020-05-01 · EMNLP 2020 11 · Emily Dinan, Angela Fan, Ledell Wu, Jason Weston 외

Machine learning models are trained to find patterns in data. NLP models can inadvertently learn socially undesirable patterns when training on gender biased text. In this work, we propose a general framework that decomp…

ClassificationGeneral Classification

Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation

2025-11-26 · Lina Conti, Dennis Fucci, Marco Gaido, Matteo Negri 외 arxiv

Unlike text, speech conveys information about the speaker, such as gender, through acoustic cues like pitch. This gives rise to modality-specific bias concerns. For example, in speech translation (ST), when translating f…

Who Finds This Voice Attractive? A Large-Scale Experiment Using In-the-Wild Data

2024-07-05 · Hitoshi Suda, Aya Watanabe, Shinnosuke Takamichi

This paper introduces CocoNut-Humoresque, an open-source large-scale speech likability corpus that includes speech segments and their per-listener likability scores. Evaluating voice likability is essential to designing …