paper-with-me

홈 › Papers

Towards an Open-Source Dutch Speech Recognition System for the Healthcare Domain

2022-06-01 · LREC 2022 6 · Cristian Tejedor-García, Berrie van der Molen, Henk van den Heuvel, Arjan van Hessen, Toine Pieters

The current largest open-source generic automatic speech recognition (ASR) system for Dutch, Kaldi_NL, does not include a domain-specific healthcare jargon in the lexicon. Commercial alternatives (e.g., Google ASR system) are also not suitable for this purpose, not only because of the lexicon issue, but they do not safeguard privacy of sensitive data sufficiently and reliably. These reasons motivate that just a small amount of medical staff employs speech technology in the Netherlands. This paper proposes an innovative ASR training method developed within the Homo Medicinalis (HoMed) project. On the semantic level it specifically targets automatic transcription of doctor-patient consultation recordings with a focus on the use of medicines. In the first stage of HoMed, the Kaldi_NL language model (LM) is fine-tuned with lists of Dutch medical terms and transcriptions of Dutch online healthcare news bulletins. Despite the acoustic challenges and linguistic complexity of the domain, we reduced the word error rate (WER) by 5.2%. The proposed method could be employed for ASR domain adaptation to other domains with sensitive and special category data. These promising results allow us to apply this methodology on highly sensitive audiovisual recordings of patient consultations at the Netherlands Institute for Health Services Research (Nivel).

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Domain AdaptationLanguage Modellingspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Multi-Graph Decoding for Code-Switching ASR

2019-06-18 · Emre Yilmaz, Samuel Cohen, Xianghu Yue, David van Leeuwen 외

In the FAME! Project, a code-switching (CS) automatic speech recognition (ASR) system for Frisian-Dutch speech is developed that can accurately transcribe the local broadcaster's bilingual archives with CS speech. This a…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Modellingspeech-recognition+1

Acoustic and Textual Data Augmentation for Improved ASR of Code-Switching Speech

2018-07-28 · Emre Yilmaz, Henk van den Heuvel, David A. van Leeuwen

In this paper, we describe several techniques for improving the acoustic and language model of an automatic speech recognition (ASR) system operating on code-switching (CS) speech. We focus on the recognition of Frisian-…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationLanguage Modeling+3

Speech Recognition Web Services for Dutch

2014-05-01 · LREC 2014 5 · Joris Pelemans, Kris Demuynck, Hugo Van hamme, Patrick Wambacq

In this paper we present 3 applications in the domain of Automatic Speech Recognition for Dutch, all of which are developed using our in-house speech recognition toolkit SPRAAK. The speech-to-text transcriber is a large …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

Code-Switching Detection with Data-Augmented Acoustic and Language Models

2018-07-28 · Yılmaz Emre, Heuvel Henk van den, van Leeuwen David A.

In this paper, we investigate the code-switching detection performance of a code-switching (CS) automatic speech recognition (ASR) system with data-augmented acoustic and language models. We focus on the recognition of F…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+3

Semi-supervised acoustic model training for speech with code-switching

2018-10-23 · Emre Yilmaz, Mitchell McLaren, Henk van den Heuvel, David A. van Leeuwen

In the FAME! project, we aim to develop an automatic speech recognition (ASR) system for Frisian-Dutch code-switching (CS) speech extracted from the archives of a local broadcaster with the ultimate goal of building a sp…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Retrievalspeaker-diarization+4