paper-with-me

Papers

Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model

2024-09-03 · Hukai Huang, Jiayan Lin, Kaidi Wang, Yishuang Li, Wenhao Guan, Lin Li, Qingyang Hong

Due to the inherent difficulty in modeling phonetic similarities across different languages, code-switching speech recognition presents a formidable challenge. This study proposes a Collaborative-MoE, a Mixture of Experts (MoE) model that leverages a collaborative mechanism among expert groups. Initially, a preceding routing network explicitly learns Language Identification (LID) tasks and selects experts based on acquired LID weights. This process ensures robust routing information to the MoE layer, mitigating interference from diverse language domains on expert network parameter updates. The LID weights are also employed to facilitate inter-group collaboration, enabling the integration of language-specific representations. Furthermore, within each language expert group, a gating network operates unsupervised to foster collaboration on attributes beyond language. Extensive experiments demonstrate the efficacy of our approach, achieving significant performance enhancements compared to alternative methods. Importantly, our method preserves the efficient inference capabilities characteristic of MoE models without necessitating additional pre-training.

📄 PDF Abstract BibTeX arXiv:2409.02050

Code (0)

등록된 구현이 없습니다.

Tasks

Language IdentificationMixture-of-Expertsspeech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

MoE 설명 없음

Similar Papers 제목 키워드 기반

Improving Code-Switching Speech Recognition with TTS Data Augmentation

2026-01-02 · Yue Heng Yeo, Yuchen Hu, Shreyas Gopal, Yizhou Peng 외 arxiv

Automatic speech recognition (ASR) for conversational code-switching speech remains challenging due to the scarcity of realistic, high-quality labeled speech data. This paper explores multilingual text-to-speech (TTS) mo…

Speech RecognitionData Augmentation

Enhancing Code-switching Speech Recognition with Interactive Language Biases

2023-09-29 · Hexin Liu, Leibny Paola Garcia, Xiangyu Zhang, Andy W. H. Khong 외

Languages usually switch within a multilingual speech signal, especially in a bilingual society. This phenomenon is referred to as code-switching (CS), making automatic speech recognition (ASR) challenging under a multil…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2

Towards End-to-End Code-Switching Speech Recognition

2018-10-31 · Ne Luo, Dongwei Jiang, Shuaijiang Zhao, Caixia Gong 외

Code-switching speech recognition has attracted an increasing interest recently, but the need for expert linguistic knowledge has always been a big issue. End-to-end automatic speech recognition (ASR) simplifies the buil…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Identificationspeech-recognition+1

Exploiting Low-Resource Code-Switching Data to Mandarin-English Speech Recognition Systems

2021-10-01 · ROCLING 2021 10 · Hou-An Lin, Chia-Ping Chen

In this paper, we investigate how to use limited code-switching data to implement a code-switching speech recognition system. We utilize the Transformer end-to-end model to develop our code switching speech recognition s…

Language ModelingLanguage ModellingMulti-Task Learningspeech-recognition+2

Gated Low-rank Adaptation for personalized Code-Switching Automatic Speech Recognition on the low-spec devices

2024-04-24 · Gwantae Kim, Bokyeung Lee, Donghyeon Kim, Hanseok Ko

In recent times, there has been a growing interest in utilizing personalized large models on low-spec devices, such as mobile and CPU-only devices. However, utilizing a personalized large model in the on-device is ineffi…

Automatic Speech RecognitionCPUparameter-efficient fine-tuningspeech-recognition+1