paper-with-me

Papers Speech Language Identification

“Speech Language Identification” 태그가 달린 논문 6편 · 필터 해제

mHuBERT-147: A Compact Multilingual HuBERT Model

2024-06-10 · Marcely Zanon Boito, Vivek Iyer, Nikolaos Lagos, Laurent Besacier 외

We present mHuBERT-147, the first general-purpose massively multilingual HuBERT speech representation model trained on 90K hours of clean, open-license data. To scale up the multi-iteration HuBERT approach, we use faiss-…

Automatic Speech Recognition (ASR)DiversitymodelSpeech Language Identification+1

FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

2022-05-25 · Alexis Conneau, Min Ma, Simran Khanuja, Yu Zhang 외

We introduce FLEURS, the Few-shot Learning Evaluation of Universal Representations of Speech benchmark. FLEURS is an n-way parallel speech dataset in 102 languages built on top of the machine translation FLoRes-101 bench…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Few-Shot LearningLanguage Identification+6

Modernizing Open-Set Speech Language Identification

2022-05-20 · Mustafa Eyceoz, Justin Lee, Homayoon Beigi

While most modern speech Language Identification methods are closed-set, we want to see if they can be modified and adapted for the open-set problem. When switching to the open-set problem, the solution gains the ability…

Language IdentificationSpeech Language Identification

Triplet Entropy Loss: Improving The Generalisation of Short Speech Language Identification Systems

2020-12-03 · Ruan van der Merwe

We present several methods to improve the generalisation of language identification (LID) systems to new speakers and to new domains. These methods involve Spectral augmentation, where spectrograms are masked in the freq…

Language IdentificationSpeech Language IdentificationSpoken language identificationSpoken Language Understanding+1

Language Identification with Deep Bottleneck Features

2018-09-18 · Zhanyu Ma, Hong Yu

In this paper we proposed an end-to-end short utterances speech language identification(SLD) approach based on a Long Short Term Memory (LSTM) neural network which is special suitable for SLD application in intelligent v…

Language IdentificationSpeech Language IdentificationTransfer Learning

Statistical Analysis of Multilingual Text Corpus and Development of Language Models

2014-05-01 · LREC 2014 5 · Shyam Sundar Agrawal, {Abhimanue}, shweta bansal, Minakshi Mahajan

This paper presents two studies, first a statistical analysis for three languages i.e. Hindi, Punjabi and Nepali and the other, development of language models for three Indian languages i.e. Indian English, Punjabi and N…

Language IdentificationLanguage ModellingSpeech Language Identification
1–6 / 6