paper-with-me

홈 › Papers

LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families

2026-03-20 · Jianan Chen, Xiaoxue Gao, Tatsuya Kawahara, Nancy F. Chen arxiv

Large language models (LLMs) have driven substantial advances in speech language models (SpeechLMs), yielding strong performance in automatic speech recognition (ASR) under high-resource conditions. However, existing benchmarks predominantly focus on high-resource languages, leaving the ASR behavior of SpeechLMs in low-resource languages insufficiently understood. This gap is critical, as practical ASR systems must reliably support low-resource languages and generalize across diverse language families, and it directly hinders the deployment of SpeechLM-based ASR in real-world multilingual scenarios. As a result, it is essential to evaluate SpeechLMs on low-resource languages to ensure their generalizability across different language families. To address this problem, we propose \textbf{LoASR-Bench}, a comprehensive benchmark designed to evaluate \textbf{lo}w-resource \textbf{a}utomatic \textbf{s}peech \textbf{r}ecognition (\textbf{ASR}) of the latest SpeechLMs across diverse language families. LoASR-Bench comprises 25 languages from 9 language families, featuring both Latin and non-Latin scripts, enabling cross-linguistic and cross-script assessment of ASR performance of current SpeechLMs. Experimental results highlight the limitations of the latest SpeechLMs in handling real-world low-resource languages.

📄 PDF Abstract BibTeX arXiv:2603.20042

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness

2026-03-30 · Jinyoung Kim, Hyeongsoo Lim, Eunseo Seo, Minho Jang 외 arxiv

Recent advances in large audio language models (LALMs) have enabled multilingual speech understanding. However, benchmarks for evaluating LALMs remain scarce for non-English languages, with Korean being one such underexp…

Instruction FollowingSpeech RecognitionQuestion Answering

SpeechR: A Benchmark for Speech Reasoning in Large Audio-Language Models

2025-08-04 · Wanqi Yang, Yanda Li, Yunchao Wei, Meng Fang 외 arxiv

Large audio-language models (LALMs) have achieved near-human performance in sentence-level transcription and emotion recognition. However, existing evaluations focus mainly on surface-level perception, leaving the capaci…

Emotion RecognitionAnswer Selection

RW-Voice-EQ Bench: A Real World Benchmark for Evaluating Voice AI Systems

2026-07-16 · David Ayllon, Alice Baird, Jeffrey Brooks, Franc Camps-Febrer 외 arxiv

Current voice AI benchmarks typically evaluate isolated capabilities such as speech intelligibility, word error rate, or text-based dialogue quality, but they rarely test whether systems harness the acoustic information …

Speech Recognition

StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control

2026-03-08 · Haishu Zhao, Aokai Hao, Yuan Ge, Zhenqiang Hong 외 arxiv

Speech language models (SLMs) have significantly extended the interactive capability of text-based Large Language Models (LLMs) by incorporating paralinguistic information. For more realistic interactive experience with …

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

2026-05-27 · Haechan Kim, Seungjun Chung, Inkyu Park, Jihoo Lee 외 arxiv

Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, SpeechLM evaluation remains heavily centered on English, limiting reliable …