paper-with-me

홈 › Papers

Cross-Corpora Language Recognition: A Preliminary Investigation with Indian Languages

2021-05-10 · Spandan Dey, Goutam Saha, Md Sahidullah

In this paper, we conduct one of the very first studies for cross-corpora performance evaluation in the spoken language identification (LID) problem. Cross-corpora evaluation was not explored much in LID research, especially for the Indian languages. We have selected three Indian spoken language corpora: IIITH-ILSC, LDC South Asian, and IITKGP-MLILSC. For each of the corpus, LID systems are trained on the state-of-the-art time-delay neural network (TDNN) based architecture with MFCC features. We observe that the LID performance degrades drastically for cross-corpora evaluation. For example, the system trained on the IIITH-ILSC corpus shows an average EER of 11.80 % and 43.34 % when evaluated with the same corpora and LDC South Asian corpora, respectively. Our preliminary analysis shows the significant differences among these corpora in terms of mismatch in the long-term average spectrum (LTAS) and signal-to-noise ratio (SNR). Subsequently, we apply different feature level compensation methods to reduce the cross-corpora acoustic mismatch. Our results indicate that these feature normalization schemes can help to achieve promising LID performance on cross-corpora experiments.

📄 PDF Abstract BibTeX arXiv:2105.04639

Code (0)

등록된 구현이 없습니다.

Tasks

Language IdentificationSpoken language identification

Similar Papers 제목 키워드 기반

Low Resource Multimodal Translation of Nepali Spoken Words into Emotion-Conditioned Sign Language Avatars

2026-05-04 · Jatin Bhusal, Salma Tamang arxiv

Sign language communication systems, that integrate emotional expression remain underexplored, particularly for low-resource languages. This pilot study presents NEST-V1 (Nepali Emotion and Speech Transformer - Version 1…

Sign Language TranslationEmotion ClassificationEmotion RecognitionSpeech Recognition

Oxymorons: a preliminary corpus investigation

2020-07-01 · WS 2020 7 · Marta La Pietra, Francesca Masini

This paper contains a preliminary corpus study of oxymorons, a figure of speech so far under-investigated in NLP-oriented research. The study resulted in a list of 376 oxymorons, identified by extracting a set of antonym…

Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices

2023-12-20 · Beltrán Labrador, Manuel Otero-Gonzalez, Alicia Lozano-Diez, Daniel Ramos 외

This paper presents VoxCeleb-ESP, a collection of pointers and timestamps to YouTube videos facilitating the creation of a novel speaker recognition dataset. VoxCeleb-ESP captures real-world scenarios, incorporating dive…

Speaker IdentificationSpeaker Recognition

Experimenting active and sequential learning in a medieval music manuscript

2025-07-21 · Sachin Sharma, Federico Simonetta, Michele Flammini arxiv

Optical Music Recognition (OMR) is a cornerstone of music digitization initiatives in cultural heritage, yet it remains limited by the scarcity of annotated data and the complexity of historical manuscripts. In this pape…

Object DetectionActive Learning

Proper Body Landmark Subset Enables More Accurate and 5X Faster Recognition of Isolated Signs in LIBRAS

2025-10-28 · Daniele L. V. dos Santos, Thiago B. Pereira, Carlos Eduardo G. R. Alves, Richard J. M. G. Tello 외 arxiv

This paper examines the feasibility of utilizing lightweight body landmark detection for recognizing isolated signs in Brazilian Sign Language (LIBRAS). Although the use of skeleton-image representation has enabled subst…