Spoken language identification
12개 벤치마크 · 논문 53편 · 이 태스크의 논문 보기 →
Benchmarks
LRE07
VoxForge European
VoxForge Commonwealth
IndicTTS
VoxForge
KALAKA-3
VOXLINGUA107
Most implemented
VoxLingua107: a Dataset for Spoken Language Recognition
Improving Multilingual ASR in the Wild Using Simple N-best Re-ranking
Unified model for code-switching speech recognition and language identification based on a concatenated tokenizer
Spoken Language Identification System for English-Mandarin Code-Switching Child-Directed Speech
Improving Spoken Language Identification with Map-Mix
Papers
Spoken Language Identification with Pre-trained Models and Margin Loss
For the speaker-controlled spoken language identification task proposed in the TidyLang Challenge 2026, this paper proposes a language identification method based on pre-trained models and margin-based losses. The propos…
Spoken language identificationGeolocation-Aware Robust Spoken Language Identification
While Self-supervised Learning (SSL) has significantly improved Spoken Language Identification (LID), existing models often struggle to consistently classify dialects and accents of the same language as a unified class. …
Spoken language identificationSelf-Supervised LearningOn the use of Performer and Agent Attention for Spoken Language Identification
One of the methods for language Identification (LID) involves deriving speech representation from pre-trained models using self-supervised learning, followed by fine-tuning the model for the LID task. State-of-the-art ap…
Language IdentificationSelf-Supervised LearningSpoken language identificationAfriHuBERT: A self-supervised speech representation model for African languages
In this work, we present AfriHuBERT, an extension of mHuBERT-147, a compact self-supervised learning (SSL) model pretrained on 147 languages. While mHuBERT-147 covered 16 African languages, we expand this to 1,226 throug…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Cross-corpusLanguage Identification+4Improving Multilingual ASR in the Wild Using Simple N-best Re-ranking
Multilingual Automatic Speech Recognition (ASR) models are typically evaluated in a setting where the ground-truth language of the speech utterance is known, however, this is often not the case for most practical setting…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language IdentificationRe-Ranking+3Exploring Spoken Language Identification Strategies for Automatic Transcription of Multilingual Broadcast and Institutional Speech
This paper addresses spoken language identification (SLI) and speech recognition of multilingual broadcast and institutional speech, real application scenarios that have been rarely addressed in the SLI literature. Obser…
Language Identificationspeaker-diarizationSpeaker Diarizationspeech-recognition+2