paper-with-me

홈 › Papers

SIGTYP 2021 Shared Task: Robust Spoken Language Identification

2021-06-07 · NAACL (SIGTYP) 2021 6 · Elizabeth Salesky, Badr M. Abdullah, Sabrina J. Mielke, Elena Klyachko, Oleg Serikov, Edoardo Ponti, Ritesh Kumar, Ryan Cotterell, Ekaterina Vylomova

While language identification is a fundamental speech and language processing task, for many languages and language families it remains a challenging task. For many low-resource and endangered languages this is in part due to resource availability: where larger datasets exist, they may be single-speaker or have different domains than desired application scenarios, demanding a need for domain and speaker-invariant language identification systems. This year's shared task on robust spoken language identification sought to investigate just this scenario: systems were to be trained on largely single-speaker speech from one domain, but evaluated on data in other domains recorded from speakers under different recording circumstances, mimicking realistic low-resource scenarios. We see that domain and speaker mismatch proves very challenging for current methods which can perform above 95% accuracy in-domain, which domain adaptation can address to some degree, but that these conditions merit further investigation to make spoken language identification accessible in many scenarios.

📄 PDF Abstract BibTeX arXiv:2106.03895

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationLanguage IdentificationSpoken language identification

Similar Papers 제목 키워드 기반

Language ID Prediction from Speech Using Self-Attentive Pooling

2021-06-01 · NAACL (SIGTYP) 2021 6 · Roman Bedyakin, Nikolay Mikhaylovskiy

This memo describes NTR-TSU submission for SIGTYP 2021 Shared Task on predicting language IDs from speech. Spoken Language Identification (LID) is an important step in a multilingual Automated Speech Recognition (ASR) sy…

Language Identificationspeech-recognitionSpeech RecognitionSpoken language identification

Language ID Prediction from Speech Using Self-Attentive Pooling and 1D-Convolutions

2021-04-24 · Roman Bedyakin, Nikolay Mikhaylovskiy

This memo describes NTR-TSU submission for SIGTYP 2021 Shared Task on predicting language IDs from speech. Spoken Language Identification (LID) is an important step in a multilingual Automated Speech Recognition (ASR) sy…

Language Identificationspeech-recognitionSpeech RecognitionSpoken language identification

Anlirika: An LSTM–CNN Flow Twister for Spoken Language Identification

2021-06-01 · NAACL (SIGTYP) 2021 6 · Andreas Scherbakov, Liam Whittle, Ritesh Kumar, Siddharth Singh 외

The paper presents Anlirika’s submission to SIGTYP 2021 Shared Task on Robust Spoken Language Identification. The task aims at building a robust system that generalizes well across different domains and speakers. The tra…

Language IdentificationSpoken language identification

KMI-Panlingua-IITKGP @SIGTYP2020: Exploring rules and hybrid systems for automatic prediction of typological features

2020-11-01 · EMNLP (SIGTYP) 2020 11 · Ritesh Kumar, Deepak Alok, Akanksha Bansal, Bornini Lahiri 외

This paper enumerates SigTyP 2020 Shared Task on the prediction of typological features as performed by the KMI-Panlingua-IITKGP team. The task entailed the prediction of missing values in a particular language, provided…

Missing Values

A ResNet-50-Based Convolutional Neural Network Model for Language ID Identification from Speech Recordings

2021-06-01 · NAACL (SIGTYP) 2021 6 · Giuseppe G. A. Celano

This paper describes the model built for the SIGTYP 2021 Shared Task aimed at identifying 18 typologically different languages from speech recordings. Mel-frequency cepstral coefficients derived from audio files are tran…