paper-with-me

Papers

Meta-Learning for improving rare word recognition in end-to-end ASR

2021-02-25 · Florian Lux, Ngoc Thang Vu

We propose a new method of generating meaningful embeddings for speech, changes to four commonly used meta learning approaches to enable them to perform keyword spotting in continuous signals and an approach of combining their outcomes into an end-to-end automatic speech recognition system to improve rare word recognition. We verify the functionality of each of our three contributions in two experiments exploring their performance for different amounts of classes (N-way) and examples per class (k-shot) in a few-shot setting. We find that the speech embeddings work well and the changes to the meta learning approaches also clearly enable them to perform continuous signal spotting. Despite the interface between keyword spotting and speech recognition being very simple, we are able to consistently improve word error rate by up to 5%.

📄 PDF Abstract BibTeX arXiv:2102.12624

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Keyword SpottingMeta-Learningspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

An analysis of language models for metaphor recognition

2020-12-01 · COLING 2020 8 · Arthur Neidlein, Philip Wiesenbach, Katja Markert

We conduct a linguistic analysis of recent metaphor recognition systems, all of which are based on language models. We show that their performance, although reaching high F-scores, has considerable gaps from a linguistic…

Word Sense Disambiguation

ed-cec: improving rare word recognition using asr postprocessing based on error detection and context-aware error correction

2023-10-08 · Jiajun He, Zekun Yang, Tomoki Toda

Automatic speech recognition (ASR) systems often encounter difficulties in accurately recognizing rare words, leading to errors that can have a negative impact on downstream tasks such as keyword spotting, intent detecti…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Intent DetectionKeyword Spotting+3

Contextual RNN-T For Open Domain ASR

2020-06-04 · Mahaveer Jain, Gil Keren, Jay Mahadeokar, Geoffrey Zweig 외

End-to-end (E2E) systems for automatic speech recognition (ASR), such as RNN Transducer (RNN-T) and Listen-Attend-Spell (LAS) blend the individual components of a traditional hybrid ASR system - acoustic model, language …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2

Multi-task Language Modeling for Improving Speech Recognition of Rare Words

2020-11-23 · Chao-Han Huck Yang, Linda Liu, Ankur Gandhe, Yile Gu 외

End-to-end automatic speech recognition (ASR) systems are increasingly popular due to their relative architectural simplicity and competitive performance. However, even though the average accuracy of these systems may be…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+3

Improving Large-scale Deep Biasing with Phoneme Features and Text-only Data in Streaming Transducer

2023-11-15 · Jin Qiu, Lu Huang, Boyu Li, Jun Zhang 외

Deep biasing for the Transducer can improve the recognition performance of rare words or contextual entities, which is essential in practical applications, especially for streaming Automatic Speech Recognition (ASR). How…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition