Internal Language Model Estimation based Language Model Fusion for Cross-Domain Code-Switching Speech Recognition
Internal Language Model Estimation (ILME) based language model (LM) fusion has been shown significantly improved recognition results over conventional shallow fusion in both intra-domain and cross-domain speech recognition tasks. In this paper, we attempt to apply our ILME method to cross-domain code-switching speech recognition (CSSR) work. Specifically, our curiosity comes from several aspects. First, we are curious about how effective the ILME-based LM fusion is for both intra-domain and cross-domain CSSR tasks. We verify this with or without merging two code-switching domains. More importantly, we train an end-to-end (E2E) speech recognition model by means of merging two monolingual data sets and observe the efficacy of the proposed ILME-based LM fusion for CSSR. Experimental results on SEAME that is from Southeast Asian and another Chinese Mainland CS data set demonstrate the effectiveness of the proposed ILME-based LM fusion method.
Code (0)
등록된 구현이 없습니다.
Tasks
Language ModelingLanguage Modellingmodelspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
Residual Language Model for End-to-end Speech Recognition
End-to-end automatic speech recognition suffers from adaptation to unknown target domain speech despite being trained with a large amount of paired audio--text data. Recent studies estimate a linguistic bias of the model…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Domain AdaptationLanguage Modeling+3Internal Language Model Estimation based Adaptive Language Model Fusion for Domain Adaptation
ASR model deployment environment is ever-changing, and the incoming speech can be switched across different domains during a session. This brings a challenge for effective domain adaptation when only target domain text d…
Domain AdaptationLanguage ModelingLanguage ModellingmodelMask The Bias: Improving Domain-Adaptive Generalization of CTC-based ASR with Internal Language Model Estimation
End-to-end ASR models trained on large amount of data tend to be implicitly biased towards language semantics of the training data. Internal language model estimation (ILME) has been proposed to mitigate this bias for au…
DecoderDomain AdaptationLanguage ModelingLanguage ModellingLibrispeech Transducer Model with Internal Language Model Prior Correction
We present our transducer model on Librispeech. We study variants to include an external language model (LM) with shallow fusion and subtract an estimated internal LM. This is justified by a Bayesian interpretation where…
Language ModelingLanguage ModellingmodelSentence+1Label-Context-Dependent Internal Language Model Estimation for CTC
Although connectionist temporal classification (CTC) has the label context independence assumption, it can still implicitly learn a context-dependent internal language model (ILM) due to modern powerful encoders. In this…
Knowledge DistillationLanguage ModelingLanguage Modelling