Improving Black-box Speech Recognition using Semantic Parsing
Speech is a natural channel for human-computer interaction in robotics and consumer applications. Natural language understanding pipelines that start with speech can have trouble recovering from speech recognition errors. Black-box automatic speech recognition (ASR) systems, built for general purpose use, are unable to take advantage of in-domain language models that could otherwise ameliorate these errors. In this work, we present a method for re-ranking black-box ASR hypotheses using an in-domain language model and semantic parser trained for a particular task. Our re-ranking method significantly improves both transcription accuracy and semantic understanding over a state-of-the-art ASR{'}s vanilla output.
Code (0)
등록된 구현이 없습니다.
Tasks
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage ModellingNatural Language UnderstandingRe-RankingSemantic Parsingspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
Semantic Parsing of Disfluent Speech
Speech disfluencies are prevalent in spontaneous speech. The rising popularity of voice assistants presents a growing need to handle naturally occurring disfluencies. Semantic parsing is a key component for understanding…
Semantic ParsingA Study on the Integration of Pipeline and E2E SLU systems for Spoken Semantic Parsing toward STOP Quality Challenge
Recently there have been efforts to introduce new benchmark tasks for spoken language understanding (SLU), like semantic parsing. In this paper, we describe our proposed spoken semantic parsing system for the quality tra…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Semantic Parsingspeech-recognition+2Growing Trees on Sounds: Assessing Strategies for End-to-End Dependency Parsing of Speech
Direct dependency parsing of the speech signal -- as opposed to parsing speech transcriptions -- has recently been proposed as a task (Pupier et al. 2022), as a way of incorporating prosodic information in the parsing sy…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Dependency Parsingspeech-recognition+1Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding
Word Error Rate (WER) has been the predominant metric used to evaluate the performance of automatic speech recognition (ASR) systems. However, WER is sometimes not a good indicator for downstream Natural Language Underst…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Intent RecognitionLanguage Modeling+14Fast and Accurate Capitalization and Punctuation for Automatic Speech Recognition Using Transformer and Chunk Merging
In recent years, studies on automatic speech recognition (ASR) have shown outstanding results that reach human parity on short speech segments. However, there are still difficulties in standardizing the output of ASR suc…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)NERPOS+4