paper-with-me

Papers

Boosting Active Learning for Speech Recognition with Noisy Pseudo-labeled Samples

2020-06-19 · Jihwan Bang, Heesu Kim, Youngjoon Yoo, Jung-Woo Ha

The cost of annotating transcriptions for large speech corpora becomes a bottleneck to maximally enjoy the potential capacity of deep neural network-based automatic speech recognition models. In this paper, we present a new training pipeline boosting the conventional active learning approach targeting label-efficient learning to resolve the mentioned problem. Existing active learning methods only focus on selecting a set of informative samples under a labeling budget. One step further, we suggest that the training efficiency can be further improved by utilizing the unlabeled samples, exceeding the labeling budget, by introducing sophisticatedly configured unsupervised loss complementing supervised loss effectively. We propose new unsupervised loss based on consistency regularization, and we configure appropriate augmentation techniques for utterances to adopt consistency regularization in the automatic speech recognition task. From the qualitative and quantitative experiments on the real-world dataset and under real-usage scenarios, we show that the proposed training pipeline can boost the efficacy of active learning approaches, thus successfully reducing a sustainable amount of human labeling cost.

📄 PDF Abstract BibTeX arXiv:2006.11021

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Self-Training for End-to-End Speech Recognition

2019-09-19 · Jacob Kahn, Ann Lee, Awni Hannun

We revisit self-training in the context of end-to-end speech recognition. We demonstrate that training with pseudo-labels can substantially improve the accuracy of a baseline model. Key to our approach are a strong basel…

DiversityLanguage ModelingLanguage ModellingPseudo Label+2

Filter and evolve: progressive pseudo label refining for semi-supervised automatic speech recognition

2022-10-28 · Zezhong Jin, Dading Zhong, Xiao Song, Zhaoyi Liu 외

Fine tuning self supervised pretrained models using pseudo labels can effectively improve speech recognition performance. But, low quality pseudo labels can misguide decision boundaries and degrade performance. We propos…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Pseudo Labelspeech-recognition+1

Alternative Pseudo-Labeling for Semi-Supervised Automatic Speech Recognition

2023-08-12 · Han Zhu, Dongji Gao, Gaofeng Cheng, Daniel Povey 외

When labeled data is insufficient, semi-supervised learning with the pseudo-labeling technique can significantly improve the performance of automatic speech recognition. However, pseudo-labels are often noisy, containing…

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

ReHear: Iterative Pseudo-Label Refinement for Semi-Supervised Speech Recognition via Audio Large Language Models

2026-02-21 · Zefang Liu, Chenyang Zhu, Sangwoo Cho, Shi-Xiong Zhang arxiv

Semi-supervised learning in automatic speech recognition (ASR) typically relies on pseudo-labeling, which often suffers from confirmation bias and error accumulation due to noisy supervision. To address this limitation, …

Speech Recognition

Interactive Feature Fusion for End-to-End Noise-Robust Speech Recognition

2021-10-11 · Yuchen Hu, Nana Hou, Chen Chen, Eng Siong Chng

Speech enhancement (SE) aims to suppress the additive noise from a noisy speech signal to improve the speech's perceptual quality and intelligibility. However, the over-suppression phenomenon in the enhanced speech might…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Robust Speech RecognitionSpeech Enhancement+2