paper-with-me

홈 › Papers

The Balancing Act: Unmasking and Alleviating ASR Biases in Portuguese

2024-02-12 · Ajinkya Kulkarni, Anna Tokareva, Rameez Qureshi, Miguel Couceiro

In the field of spoken language understanding, systems like Whisper and Multilingual Massive Speech (MMS) have shown state-of-the-art performances. This study is dedicated to a comprehensive exploration of the Whisper and MMS systems, with a focus on assessing biases in automatic speech recognition (ASR) inherent to casual conversation speech specific to the Portuguese language. Our investigation encompasses various categories, including gender, age, skin tone color, and geo-location. Alongside traditional ASR evaluation metrics such as Word Error Rate (WER), we have incorporated p-value statistical significance for gender bias analysis. Furthermore, we extensively examine the impact of data distribution and empirically show that oversampling techniques alleviate such stereotypical biases. This research represents a pioneering effort in quantifying biases in the Portuguese language context through the application of MMS and Whisper, contributing to a better understanding of ASR systems' performance in multilingual settings.

📄 PDF Abstract BibTeX arXiv:2402.07513

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionSpoken Language Understanding

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Step-by-Step Unmasking for Parameter-Efficient Fine-tuning of Large Language Models

2024-08-26 · Aradhye Agarwal, Suhas K Ramesh, Ayan Sengupta, Tanmoy Chakraborty

Fine-tuning large language models (LLMs) on downstream tasks requires substantial computational resources. A class of parameter-efficient fine-tuning (PEFT) aims to mitigate these computational challenges by selectively …

Computational EfficiencyNatural Language Understandingparameter-efficient fine-tuning

Performance in a dialectal profiling task of LLMs for varieties of Brazilian Portuguese

2024-10-14 · Raquel Meister Ko Freitag, Túlio Sousa de Gois

Different of biases are reproduced in LLM-generated responses, including dialectal biases. A study based on prompt engineering was carried out to uncover how LLMs discriminate varieties of Brazilian Portuguese, specifica…

Prompt Engineering

Enhancing Portuguese Variety Identification with Cross-Domain Approaches

2025-02-20 · Hugo Sousa, Rúben Almeida, Purificação Silvano, Inês Cantante 외

Recent advances in natural language processing have raised expectations for generative models to produce coherent text across diverse language varieties. In the particular case of the Portuguese language, the predominanc…

CLIP the Bias: How Useful is Balancing Data in Multimodal Learning?

2024-03-07 · Ibrahim Alabdulmohsin, Xiao Wang, Andreas Steiner, Priya Goyal 외

We study the effectiveness of data-balancing for mitigating biases in contrastive language-image pretraining (CLIP), identifying areas of strength and limitation. First, we reaffirm prior conclusions that CLIP models can…

Image to textImage-to-Text RetrievalRetrievalText Retrieval

Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation

2024-01-25 · Yifan Yang, Xiaoyu Liu, Qiao Jin, Furong Huang 외

Large language models like GPT-3.5-turbo and GPT-4 hold promise for healthcare professionals, but they may inadvertently inherit biases during their training, potentially affecting their utility in medical applications. …

Medical Report Generation