paper-with-me

Papers

Adversarial Training for Multilingual Acoustic Modeling

2019-06-17 · Ke Hu, Hasim Sak, Hank Liao

Multilingual training has been shown to improve acoustic modeling performance by sharing and transferring knowledge in modeling different languages. Knowledge sharing is usually achieved by using common lower-level layers for different languages in a deep neural network. Recently, the domain adversarial network was proposed to reduce domain mismatch of training data and learn domain-invariant features. It is thus worth exploring whether adversarial training can further promote knowledge sharing in multilingual models. In this work, we apply the domain adversarial network to encourage the shared layers of a multilingual model to learn language-invariant features. Bidirectional Long Short-Term Memory (LSTM) recurrent neural networks (RNN) are used as building blocks. We show that shared layers learned this way contain less language identification information and lead to better performance. In an automatic speech recognition task for seven languages, the resultant acoustic model improves the word error rate (WER) of the multilingual model by 4% relative on average, and the monolingual models by 10%.

📄 PDF Abstract BibTeX arXiv:1906.07093

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Identificationspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Automatic Speech Recognition for Uyghur through Multilingual Acoustic Modeling

2020-05-01 · LREC 2020 5 · Ayimunishagu Abulimiti, Tanja Schultz

Low-resource languages suffer from lower performance of Automatic Speech Recognition (ASR) system due to the lack of data. As a common approach, multilingual training has been applied to achieve more context coverage and…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Neural Language Codes for Multilingual Acoustic Models

2018-07-05 · Markus Müller, Sebastian Stüker, Alex Waibel

Multilingual Speech Recognition is one of the most costly AI problems, because each language (7,000+) and even different accents require their own acoustic models to obtain best recognition performance. Even though they …

speech-recognitionSpeech Recognition

Unsupervised Pattern Discovery from Thematic Speech Archives Based on Multilingual Bottleneck Features

2020-11-03 · Man-Ling Sung, Siyuan Feng, Tan Lee

The present study tackles the problem of automatically discovering spoken keywords from untranscribed audio archives without requiring word-by-word speech transcription by automatic speech recognition (ASR) technology. T…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Clusteringspeech-recognition+1

Multilingual Graphemic Hybrid ASR with Massive Data Augmentation

2019-09-14 · LREC 2020 5 · Chunxi Liu, Qiaochu Zhang, Xiaohui Zhang, Kritika Singh 외

Towards developing high-performing ASR for low-resource languages, approaches to address the lack of resources are to make use of data from multiple languages, and to augment the training data by creating acoustic variat…

Data Augmentation

How Phonotactics Affect Multilingual and Zero-shot ASR Performance

2020-10-22 · Siyuan Feng, Piotr Żelasko, Laureano Moro-Velázquez, Ali Abavisani 외

The idea of combining multiple languages' recordings to train a single automatic speech recognition (ASR) model brings the promise of the emergence of universal speech representation. Recently, a Transformer encoder-deco…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderLanguage Modelling+2