paper-with-me

홈 › Papers

ViMedCSS: A Vietnamese Medical Code-Switching Speech Dataset & Benchmark

2026-02-13 · Tung X. Nguyen, Nhu Vo, Giang-Son Nguyen, Duy Mai Hoang, Chien Dinh Huynh, Inigo Jauregi Unanue, Massimo Piccardi, Wray Buntine, Dung D. Le arxiv

Code-switching (CS), which is when Vietnamese speech uses English words like drug names or procedures, is a common phenomenon in Vietnamese medical communication. This creates challenges for Automatic Speech Recognition (ASR) systems, especially in low-resource languages like Vietnamese. Current most ASR systems struggle to recognize correctly English medical terms within Vietnamese sentences, and no benchmark addresses this challenge. In this paper, we construct a 34-hour Vietnamese Medical Code-Switching Speech dataset (ViMedCSS) containing 16,576 utterances. Each utterance includes at least one English medical term drawn from a curated bilingual lexicon covering five medical topics. Using this dataset, we evaluate several state-of-the-art ASR models and examine different specific fine-tuning strategies for improving medical term recognition to investigate the best approach to solve in the dataset. Experimental results show that Vietnamese-optimized models perform better on general segments, while multilingual pretraining helps capture English insertions. The combination of both approaches yields the best balance between overall and code-switched accuracy. This work provides the first benchmark for Vietnamese medical code-switching and offers insights into effective domain adaptation for low-resource, multilingual ASR systems.

📄 PDF Abstract BibTeX arXiv:2602.12911

Code (0)

등록된 구현이 없습니다.

Tasks

Speech RecognitionDomain Adaptation

Similar Papers 제목 키워드 기반

CanVEC - the Canberra Vietnamese-English Code-switching Natural Speech Corpus

2020-05-01 · LREC 2020 5 · Li Nguyen, Christopher Bryant

This paper introduces the Canberra Vietnamese-English Code-switching corpus (CanVEC), an original corpus of natural mixed speech that we semi-automatically annotated with language information, part of speech (POS) tags a…

POS

Contrastive Training with LLM-generated Near-Misses for Robust Code-Switching Speech Recognition

2026-06-05 · Tung X. Nguyen, Hieu Minh Truong, Giang Son Nguyen, Nhu Vo 외 arxiv

Code-switching (CS), the alternation between multiple languages within a single utterance, remains challenging for Automatic Speech Recognition (ASR). To address this issue, we propose a Point-of-Interest (POI)-aware con…

Speech Recognition

TSPC: A Two-Stage Phoneme-Centric Architecture for code-switching Vietnamese-English Speech Recognition

2025-09-07 · Tran Nguyen Anh, Truong Dinh Dung, Vo Van Nam, Minh N. H. Nguyen arxiv

Code-switching (CS) presents a significant challenge for general Auto-Speech Recognition (ASR) systems. Existing methods often fail to capture the sub tle phonological shifts inherent in CS scenarios. The challenge is pa…

Speech Recognition

VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain

2024-04-08 · Khai Le-Duc

Due to privacy restrictions, there's a shortage of publicly available speech recognition datasets in the medical domain. In this work, we present VietMed - a Vietnamese speech recognition dataset in the medical domain co…

Language ModellingSpeech RecognitionUnsupervised Pre-trainingVietnamese Speech Recognition

AdaCS: Adaptive Normalization for Enhanced Code-Switching ASR

2025-01-13 · The Chuong Chu, Vu Tuan Dat Pham, Kien Dao, Hoang Nguyen 외

Intra-sentential code-switching (CS) refers to the alternation between languages that happens within a single utterance and is a significant challenge for Automatic Speech Recognition (ASR) systems. For example, when a V…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Decoderspeech-recognition+1