paper-with-me

Papers

Svarah: Evaluating English ASR Systems on Indian Accents

2023-05-25 · Tahir Javed, Sakshi Joshi, Vignesh Nagarajan, Sai Sundaresan, Janki Nawale, Abhigyan Raman, Kaushal Bhogale, Pratyush Kumar, Mitesh M. Khapra

India is the second largest English-speaking country in the world with a speaker base of roughly 130 million. Thus, it is imperative that automatic speech recognition (ASR) systems for English should be evaluated on Indian accents. Unfortunately, Indian speakers find a very poor representation in existing English ASR benchmarks such as LibriSpeech, Switchboard, Speech Accent Archive, etc. In this work, we address this gap by creating Svarah, a benchmark that contains 9.6 hours of transcribed English audio from 117 speakers across 65 geographic locations throughout India, resulting in a diverse range of accents. Svarah comprises both read speech and spontaneous conversational data, covering various domains, such as history, culture, tourism, etc., ensuring a diverse vocabulary. We evaluate 6 open source ASR models and 2 commercial ASR systems on Svarah and show that there is clear scope for improvement on Indian accents. Svarah as well as all our code will be publicly available.

📄 PDF Abstract BibTeX arXiv:2305.15760

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

An Investigation of Indian Native Language Phonemic Influences on L2 English Pronunciations

2022-12-19 · Shelly Jain, Priyanshi Pal, Anil Vuppala, Prasanta Ghosh 외

Speech systems are sensitive to accent variations. This is especially challenging in the Indian context, with an abundance of languages but a dearth of linguistic studies characterising pronunciation variations. The grow…

AccentDB: A Database of Non-Native English Accents to Assist Neural Speech Recognition

2020-05-16 · LREC 2020 5 · Afroz Ahamad, Ankit Anand, Pranesh Bhargava

Modern Automatic Speech Recognition (ASR) technology has evolved to identify the speech spoken by native speakers of a language very well. However, identification of the speech spoken by non-native speakers continues to …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Deep Speech Based End-to-End Automated Speech Recognition (ASR) for Indian-English Accents

2022-04-03 · Priyank Dubey, Bilal Shah

Automated Speech Recognition (ASR) is an interdisciplinary application of computer science and linguistics that enable us to derive the transcription from the uttered speech waveform. It finds several applications in Mil…

speech-recognitionSpeech RecognitionSpeech-to-TextTransfer Learning

Discovering Canonical Indian English Accents: A Crowdsourcing-based Approach

2018-05-01 · LREC 2018 5 · Sunayana Sitaram, Varun Manjunath, Varun Bharadwaj, Monojit Choudhury 외
Automatic Speech Recognition (ASR)Speech Recognition

AppTek Call-Center Dialogues: A Multi-Accent Long-Form Benchmark for English ASR

2026-04-30 · Eugen Beck, Sarah Beranek, Uma Moothiringote, Daniel Mann 외 arxiv

Evaluating English ASR systems for conversational AI applications remains difficult, as many publicly available corpora are either pre-segmented into short segments, consist of read or prepared speech, or lack explicit d…