paper-with-me

Papers

Investigating data partitioning strategies for crosslinguistic low-resource ASR evaluation

2022-08-26 · Zoey Liu, Justin Spence, Emily Prud'hommeaux

Many automatic speech recognition (ASR) data sets include a single pre-defined test set consisting of one or more speakers whose speech never appears in the training set. This "hold-speaker(s)-out" data partitioning strategy, however, may not be ideal for data sets in which the number of speakers is very small. This study investigates ten different data split methods for five languages with minimal ASR training resources. We find that (1) model performance varies greatly depending on which speaker is selected for testing; (2) the average word error rate (WER) across all held-out speakers is comparable not only to the average WER over multiple random splits but also to any given individual random split; (3) WER is also generally comparable when the data is split heuristically or adversarially; (4) utterance duration and intensity are comparatively more predictive factors of variability regardless of the data split. These results suggest that the widely used hold-speakers-out approach to ASR data partitioning can yield results that do not reflect model performance on unseen data or speakers. Random splits can yield more reliable and generalizable estimates when facing data sparsity.

📄 PDF Abstract BibTeX arXiv:2208.12888

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

The Effects of Partitioning Strategies on Energy Consumption in Distributed CNN Inference at The Edge

2022-10-15 · Erqian Tang, Xiaotian Guo, Todor Stefanov

Nowadays, many AI applications utilizing resource-constrained edge devices (e.g., small mobile robots, tiny IoT devices, etc.) require Convolutional Neural Network (CNN) inference on a distributed system at the edge due …

The taggedPBC: Annotating a massive parallel corpus for crosslinguistic investigations

2025-05-18 · Hiram Ring

Existing datasets available for crosslinguistic investigations have tended to focus on large amounts of data for a small group of languages or a small amount of data for a large number of languages. This means that claim…

POS

Language Models as Artificial Learners: Investigating Crosslinguistic Influence

2026-01-29 · Abderrahmane Issam, Yusuf Can Semerci, Jan Scholtes, Gerasimos Spanakis arxiv

Despite the centrality of crosslinguistic influence (CLI) to bilingualism research, human studies often yield conflicting results due to inherent experimental variance. We address these inconsistencies by using language …

Quantifying and Reducing Speaker Heterogeneity within the Common Voice Corpus for Phonetic Analysis

2025-05-31 · Miao Zhang, Aref Farhadipour, Annie Baker, Jiachen Ma 외

With its crosslinguistic and cross-speaker diversity, the Mozilla Common Voice Corpus (CV) has been a valuable resource for multilingual speech technology and holds tremendous potential for research in crosslinguistic ph…

Diversity

Data-driven Model Generalizability in Crosslinguistic Low-resource Morphological Segmentation

2022-01-05 · Zoey Liu, Emily Prud'hommeaux

Common designs of model evaluation typically focus on monolingual settings, where different models are compared according to their performance on a single data set that is assumed to be representative of all possible dat…