paper-with-me

Spoken language identification

12개 벤치마크 · 논문 53편 · 이 태스크의 논문 보기 →

Benchmarks

LRE07

결과 27개

VoxForge European

결과 15개

VoxForge Commonwealth

결과 12개

IndicTTS

결과 9개

VoxForge

결과 9개

KALAKA-3

결과 6개

VOXLINGUA107

결과 6개

Most implemented

Papers

Spoken Language Identification with Pre-trained Models and Margin Loss

2026-05-03 · Zhihua Fang, Liang He, Weiwu Jiang arxiv

For the speaker-controlled spoken language identification task proposed in the TidyLang Challenge 2026, this paper proposes a language identification method based on pre-trained models and margin-based losses. The propos…

Spoken language identification

Geolocation-Aware Robust Spoken Language Identification

2025-08-23 · Qingzheng Wang, Hye-jin Shim, Jiancheng Sun, Shinji Watanabe arxiv

While Self-supervised Learning (SSL) has significantly improved Spoken Language Identification (LID), existing models often struggle to consistently classify dialects and accents of the same language as a unified class. …

Spoken language identificationSelf-Supervised Learning

On the use of Performer and Agent Attention for Spoken Language Identification

2025-02-09 · Jitendra Kumar dhiman, Jainag Ambati

One of the methods for language Identification (LID) involves deriving speech representation from pre-trained models using self-supervised learning, followed by fine-tuning the model for the LID task. State-of-the-art ap…

Language IdentificationSelf-Supervised LearningSpoken language identification

AfriHuBERT: A self-supervised speech representation model for African languages

2024-09-30 · Jesujoba O. Alabi, Xuechen Liu, Dietrich Klakow, Junichi Yamagishi

In this work, we present AfriHuBERT, an extension of mHuBERT-147, a compact self-supervised learning (SSL) model pretrained on 147 languages. While mHuBERT-147 covered 16 African languages, we expand this to 1,226 throug…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Cross-corpusLanguage Identification+4

Improving Multilingual ASR in the Wild Using Simple N-best Re-ranking

2024-09-27 · Brian Yan, Vineel Pratap, Shinji Watanabe, Michael Auli

Multilingual Automatic Speech Recognition (ASR) models are typically evaluated in a setting where the ground-truth language of the speech utterance is known, however, this is often not the case for most practical setting…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language IdentificationRe-Ranking+3

Exploring Spoken Language Identification Strategies for Automatic Transcription of Multilingual Broadcast and Institutional Speech

2024-06-13 · Martina Valente, Fabio Brugnara, Giovanni Morrone, Enrico Zovato 외

This paper addresses spoken language identification (SLI) and speech recognition of multilingual broadcast and institutional speech, real application scenarios that have been rarely addressed in the SLI literature. Obser…

Language Identificationspeaker-diarizationSpeaker Diarizationspeech-recognition+2

전체 53편 보기 →