paper-with-me

Papers

RAS: a Reliability Oriented Metric for Automatic Speech Recognition

2026-04-27 · Wenbin Huang, Yuhang Qiu, Bohan Li, Yiwei Guo, Jing Peng, Hankun Wang, Xie Chen, Kai Yu arxiv

Automatic speech recognition systems often produce confident yet incorrect transcriptions under noisy or ambiguous conditions, which can be misleading for both users and downstream applications. Standard evaluation based on Word Error Rate focuses solely on accuracy and fails to capture transcription reliability. We introduce an abstention-aware transcription framework that enables ASR models to explicitly abstain from uncertain segments. To evaluate reliability under abstention, we propose RAS, a reliability-oriented metric that balances transcription informativeness and error aversion, with its trade-off parameter calibrated by human preference. We then train an abstention-aware ASR model through supervised bootstrapping followed by reinforcement learning. Our experiments demonstrate substantial improvements in transcription reliability while maintaining competitive accuracy.

📄 PDF Abstract BibTeX arXiv:2604.24278

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningSpeech Recognition

Similar Papers 제목 키워드 기반

HATS: An Open data set Integrating Human Perception Applied to the Evaluation of Automatic Speech Recognition Metrics

2026-04-30 · Thibault Bañeras Roux, Jane Wottawa, Mickael Rouvier, Teva Merlin 외 arxiv

Conventionally, Automatic Speech Recognition (ASR) systems are evaluated on their ability to correctly recognize each word contained in a speech signal. In this context, the word error rate (WER) metric is the reference …

Speech Recognition

Evaluating Automatic Speech Recognition in an Incremental Setting

2023-02-23 · Ryan Whetten, Mir Tahsin Imtiaz, Casey Kennington

The increasing reliability of automatic speech recognition has proliferated its everyday use. However, for research purposes, it is often unclear which model one should choose for a task, particularly if there is a requi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

ClovaCall: Korean Goal-Oriented Dialog Speech Corpus for Automatic Speech Recognition of Contact Centers

2020-04-20 · Jung-Woo Ha, Kihyun Nam, Jingu Kang, Sang-Woo Lee 외

Automatic speech recognition (ASR) via call is essential for various applications, including AI for contact center (AICC) services. Despite the advancement of ASR, however, most publicly available call-based speech corpo…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Goal-Oriented DialogOpen-Domain Dialog+3

MoLE : Mixture of Language Experts for Multi-Lingual Automatic Speech Recognition

2023-02-27 · Yoohwan Kwon, Soo-Whan Chung

Multi-lingual speech recognition aims to distinguish linguistic expressions in different languages and integrate acoustic processing simultaneously. In contrast, current multi-lingual speech recognition research follows …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

ROSE: A Recognition-Oriented Speech Enhancement Framework in Air Traffic Control Using Multi-Objective Learning

2023-12-11 · Xincheng Yu, Dongyue Guo, Jianwei Zhang, Yi Lin

Radio speech echo is a specific phenomenon in the air traffic control (ATC) domain, which degrades speech quality and further impacts automatic speech recognition (ASR) accuracy. In this work, a time-domain recognition-o…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+1