Papers Speech Language Identification
“Speech Language Identification” 태그가 달린 논문 6편 · 필터 해제
mHuBERT-147: A Compact Multilingual HuBERT Model
We present mHuBERT-147, the first general-purpose massively multilingual HuBERT speech representation model trained on 90K hours of clean, open-license data. To scale up the multi-iteration HuBERT approach, we use faiss-…
Automatic Speech Recognition (ASR)DiversitymodelSpeech Language Identification+1FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech
We introduce FLEURS, the Few-shot Learning Evaluation of Universal Representations of Speech benchmark. FLEURS is an n-way parallel speech dataset in 102 languages built on top of the machine translation FLoRes-101 bench…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Few-Shot LearningLanguage Identification+6Modernizing Open-Set Speech Language Identification
While most modern speech Language Identification methods are closed-set, we want to see if they can be modified and adapted for the open-set problem. When switching to the open-set problem, the solution gains the ability…
Language IdentificationSpeech Language IdentificationTriplet Entropy Loss: Improving The Generalisation of Short Speech Language Identification Systems
We present several methods to improve the generalisation of language identification (LID) systems to new speakers and to new domains. These methods involve Spectral augmentation, where spectrograms are masked in the freq…
Language IdentificationSpeech Language IdentificationSpoken language identificationSpoken Language Understanding+1Language Identification with Deep Bottleneck Features
In this paper we proposed an end-to-end short utterances speech language identification(SLD) approach based on a Long Short Term Memory (LSTM) neural network which is special suitable for SLD application in intelligent v…
Language IdentificationSpeech Language IdentificationTransfer LearningStatistical Analysis of Multilingual Text Corpus and Development of Language Models
This paper presents two studies, first a statistical analysis for three languages i.e. Hindi, Punjabi and Nepali and the other, development of language models for three Indian languages i.e. Indian English, Punjabi and N…
Language IdentificationLanguage ModellingSpeech Language Identification