paper-with-me

홈 › Papers

Cross-Domain Adaptation of Spoken Language Identification for Related Languages: The Curious Case of Slavic Languages

2020-08-02 · Badr M. Abdullah, Tania Avgustinova, Bernd Möbius, Dietrich Klakow

State-of-the-art spoken language identification (LID) systems, which are based on end-to-end deep neural networks, have shown remarkable success not only in discriminating between distant languages but also between closely-related languages or even different spoken varieties of the same language. However, it is still unclear to what extent neural LID models generalize to speech samples with different acoustic conditions due to domain shift. In this paper, we present a set of experiments to investigate the impact of domain mismatch on the performance of neural LID systems for a subset of six Slavic languages across two domains (read speech and radio broadcast) and examine two low-level signal descriptors (spectral and cepstral features) for this task. Our experiments show that (1) out-of-domain speech samples severely hinder the performance of neural LID models, and (2) while both spectral and cepstral features show comparable performance within-domain, spectral features show more robustness under domain mismatch. Moreover, we apply unsupervised domain adaptation to minimize the discrepancy between the two domains in our study. We achieve relative accuracy improvements that range from 9% to 77% depending on the diversity of acoustic conditions in the source domain.

📄 PDF Abstract BibTeX arXiv:2008.00545

Code (1)

uds-lsv/da-lang-id 공식 구현 pytorch

Tasks

DiversityDomain AdaptationLanguage IdentificationSpoken language identificationUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

SIGTYP 2021 Shared Task: Robust Spoken Language Identification

2021-06-07 · NAACL (SIGTYP) 2021 6 · Elizabeth Salesky, Badr M. Abdullah, Sabrina J. Mielke, Elena Klyachko 외

While language identification is a fundamental speech and language processing task, for many languages and language families it remains a challenging task. For many low-resource and endangered languages this is in part d…

Domain AdaptationLanguage IdentificationSpoken language identification

Unsupervised neural adaptation model based on optimal transport for spoken language identification

2020-12-24 · Xugang Lu, Peng Shen, Yu Tsao, Hisashi Kawai

Due to the mismatch of statistical distributions of acoustic speech between training and testing sets, the performance of spoken language identification (SLID) could be drastically degraded. In this paper, we propose an …

Language IdentificationSpoken language identification

Deep learning-based end-to-end spoken language identification system for domain-mismatched scenario

2022-06-01 · LREC 2022 6 · Woohyun Kang, Md Jahangir Alam, Abderrahim Fathan

Domain mismatch is a critical issue when it comes to spoken language identification. To overcome the domain mismatch problem, we have applied several architectures and deep learning strategies which have shown good resul…

Language IdentificationSpeaker VerificationSpoken language identification

Partial Coupling of Optimal Transport for Spoken Language Identification

2022-03-31 · Xugang Lu, Peng Shen, Yu Tsao, Hisashi Kawai

In order to reduce domain discrepancy to improve the performance of cross-domain spoken language identification (SLID) system, as an unsupervised domain adaptation (UDA) method, we have proposed a joint distribution alig…

Domain AdaptationLanguage IdentificationSpoken language identificationUnsupervised Domain Adaptation

English-Indonesian Neural Machine Translation for Spoken Language Domains

2019-07-01 · ACL 2019 7 · Meisyarah Dwiastuti

In this work, we conduct a study on Neural Machine Translation (NMT) for English-Indonesian (EN-ID) and Indonesian-English (ID-EN). We focus on spoken language domains, namely colloquial and speech languages. We build NM…

Domain AdaptationMachine TranslationNMTTranslation