paper-with-me

Papers

Cantonese Automatic Speech Recognition Using Transfer Learning from Mandarin

2019-11-21 · Bryan Li, Xinyue Wang, Homayoon Beigi

We propose a system to develop a basic automatic speech recognizer(ASR) for Cantonese, a low-resource language, through transfer learning of Mandarin, a high-resource language. We take a time-delayed neural network trained on Mandarin, and perform weight transfer of several layers to a newly initialized model for Cantonese. We experiment with the number of layers transferred, their learning rates, and pretraining i-vectors. Key findings are that this approach allows for quicker training time with less data. We find that for every epoch, log-probability is smaller for transfer learning models compared to a Cantonese-only model. The transfer learning models show slight improvement in CER.

📄 PDF Abstract BibTeX arXiv:1911.09271

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

Cross-Lingual Cross-Age Group Adaptation for Low-Resource Elderly Speech Emotion Recognition

2023-06-26 · Samuel Cahyawijaya, Holy Lovenia, Willy Chung, Rita Frieske 외

Speech emotion recognition plays a crucial role in human-computer interactions. However, most speech emotion recognition research is biased toward English-speaking adults, which hinders its applicability to other demogra…

Data AugmentationEmotion RecognitionSpeech Emotion Recognition

Improving Rare Words Recognition through Homophone Extension and Unified Writing for Low-resource Cantonese Speech Recognition

2023-02-02 · Holam Chung, Junan Li, Pengfei Liu1, Wai-Kim Leung 외

Homophone characters are common in tonal syllable-based languages, such as Mandarin and Cantonese. The data-intensive end-to-end Automatic Speech Recognition (ASR) systems are more likely to mis-recognize homophone chara…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2

\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding

2025-10-23 · Junghyun Min, York Hay Ng, Sophia Chan, Helena Shunhua Zhao 외 arxiv

Cantonese, although spoken by millions, remains under-resourced due to policy and diglossia. To address this scarcity of evaluation frameworks for Cantonese, we introduce \textsc{\textbf{CantoNLU}}, a benchmark for Canto…

Natural Language UnderstandingNatural Language InferenceWord Sense DisambiguationLinguistic Acceptability

Pronunciation Modeling of Foreign Words for Mandarin ASR by Considering the Effect of Language Transfer

2022-10-07 · Lei Wang, Rong Tong

One of the challenges in automatic speech recognition is foreign words recognition. It is observed that a speaker's pronunciation of a foreign word is influenced by his native language knowledge, and such phenomenon is k…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Automatic Speech Recognition Datasets in Cantonese: A Survey and New Dataset

2022-01-07 · LREC 2022 6 · Tiezheng Yu, Rita Frieske, Peng Xu, Samuel Cahyawijaya 외

Automatic speech recognition (ASR) on low resource languages improves the access of linguistic minorities to technological advantages provided by artificial intelligence (AI). In this paper, we address the problem of dat…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Cultural Vocal Bursts Intensity PredictionPhilosophy+2