paper-with-me

홈 › Papers

ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages

2025-11-30 · Subham Kumar, Prakrithi Shivaprakash, Abhishek Manoharan, Astut Kurariya, Diptadhi Mukherjee, Lekhansh Shukla, Animesh Mukherjee, Prabhat Chand, Pratima Murthy arxiv

Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diverse Indian healthcare contexts remains largely unknown. In this study, we conduct the first systematic audit of ASR performance on real world clinical interview data spanning Kannada, Hindi, and Indian English, comparing leading models including Indic Whisper, Whisper, Sarvam, Google speech to text, Gemma3n, Omnilingual, Vaani, and Gemini. We evaluate transcription accuracy across languages, speakers, and demographic subgroups, with a particular focus on error patterns affecting patients vs. clinicians and gender based or intersectional disparities. Our results reveal substantial variability across models and languages, with some systems performing competitively on Indian English but failing on code mixed or vernacular speech. We also uncover systematic performance gaps tied to speaker role and gender, raising concerns about equitable deployment in clinical settings. By providing a comprehensive multilingual benchmark and fairness analysis, our work highlights the need for culturally and demographically inclusive ASR development for healthcare ecosystem in India.

📄 PDF Abstract BibTeX arXiv:2512.10967

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

Stethoscope-guided Supervised Contrastive Learning for Cross-domain Adaptation on Respiratory Sound Classification

2023-12-15 · June-Woo Kim, Sangmin Bae, Won-Yang Cho, Byungjo Lee 외

Despite the remarkable advances in deep learning technology, achieving satisfactory performance in lung sound classification remains a challenge due to the scarcity of available data. Moreover, the respiratory sound samp…

Audio ClassificationContrastive LearningDomain AdaptationLung Sound Classification+1

Evaluating Gender Bias in Speech Translation

2020-10-27 · LREC 2022 6 · Marta R. Costa-jussà, Christine Basta, Gerard I. Gállego

The scientific community is increasingly aware of the necessity to embrace pluralism and consistently represent major and minor social groups. Currently, there are no standard evaluation techniques for different types of…

Translation

Blind Source Separation in Biomedical Signals Using Variational Methods

2025-06-23 · Yasaman Torabi, Shahram Shirani, James P. Reilly

This study introduces a novel unsupervised approach for separating overlapping heart and lung sounds using variational autoencoders (VAEs). In clinical settings, these sounds often interfere with each other, making manua…

blind source separationDecoderDiagnostic

Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes

2023-09-12 · Constantino Álvarez Casado, Manuel Lage Cañellas, Matteo Pedone, Xiaoting Wu 외

Respiratory diseases remain a leading cause of mortality worldwide, highlighting the need for faster and more accurate diagnostic tools. This work presents a novel approach leveraging digital stethoscope technology for a…

Binary ClassificationClassificationDiagnosticGender Classification+1

Spoken Stereoset: On Evaluating Social Bias Toward Speaker in Speech Large Language Models

2024-08-14 · Yi-Cheng Lin, Wei-Chih Chen, Hung-Yi Lee

Warning: This paper may contain texts with uncomfortable content. Large Language Models (LLMs) have achieved remarkable performance in various tasks, including those involving multimodal data like speech. However, these …