paper-with-me

홈 › Papers

Learning pronunciation from a foreign language in speech synthesis networks

2018-10-22 · Anonymous

Although there are more than 65,000 languages in the world, the pronunciations of many phonemes sound similar across the languages. When people learn a foreign language, their pronunciation often reflect their native language's characteristics. That motivates us to investigate how the speech synthesis network learns the pronunciation when multi-lingual dataset is given. In this study, we train the speech synthesis network bilingually in English and Korean, and analyze how the network learns the relations of phoneme pronunciation between the languages. Our experimental result shows that the learned phoneme embedding vectors are located closer if their pronunciations are similar across the languages. Based on the result, we also show that it is possible to train networks that synthesize English speaker's Korean speech and vice versa. In another experiment, we train the network with limited amount of English dataset and large Korean dataset, and analyze the required amount of dataset to train a resource-poor language with the help of resource-rich languages.

📄 PDF Abstract BibTeX

Code (1)

Kyubyong/g2p 공식 구현 tf

Tasks

Speech Synthesis

Similar Papers 제목 키워드 기반

Learning pronunciation from a foreign language in speech synthesis networks

2018-11-23 · Young-Gun Lee, Suwon Shon, Taesu Kim

Although there are more than 6,500 languages in the world, the pronunciations of many phonemes sound similar across the languages. When people learn a foreign language, their pronunciation often reflects their native lan…

Speech Synthesis

Improving Cross-lingual Speech Synthesis with Triplet Training Scheme

2022-02-22 · Jianhao Ye, Hongbin Zhou, Zhiba Su, Wendi He 외

Recent advances in cross-lingual text-to-speech (TTS) made it possible to synthesize speech in a language foreign to a monolingual speaker. However, there is still a large gap between the pronunciation of generated cross…

Speech Synthesistext-to-speechText to SpeechTriplet

Exploring Retraining-Free Speech Recognition for Intra-sentential Code-Switching

2021-08-27 · Zhen Huang, Xiaodan Zhuang, Daben Liu, Xiaoqiang Xiao 외

In this paper, we present our initial efforts for building a code-switching (CS) speech recognition system leveraging existing acoustic models (AMs) and language models (LMs), i.e., no training required, and specifically…

Language ModelingLanguage Modellingspeech-recognitionSpeech Recognition

Pronunciation Modeling of Foreign Words for Mandarin ASR by Considering the Effect of Language Transfer

2022-10-07 · Lei Wang, Rong Tong

One of the challenges in automatic speech recognition is foreign words recognition. It is observed that a speaker's pronunciation of a foreign word is influenced by his native language knowledge, and such phenomenon is k…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Design of a novel Korean learning application for efficient pronunciation correction

2022-05-04 · Minjong Cheon, Minseon Kim, Hanseon Joo

The Korean wave, which denotes the global popularity of South Korea's cultural economy, contributes to the increasing demand for the Korean language. However, as there does not exist any application for foreigners to lea…

Sentencespeech-recognitionSpeech RecognitionSpeech-to-Text