Investigating Speech Recognition for Improving Predictive AAC
Making good letter or word predictions can help accelerate the communication of users of high-tech AAC devices. This is particularly important for real-time person-to-person conversations. We investigate whether per forming speech recognition on the speaking-side of a conversation can improve language model based predictions. We compare the accuracy of three plausible microphone deployment options and the accuracy of two commercial speech recognition engines (Google and IBM Watson). We found that despite recognition word error rates of 7-16{\%}, our ensemble of N-gram and recurrent neural network language models made predictions nearly as good as when they used the reference transcripts.
Code (0)
등록된 구현이 없습니다.
Tasks
Language ModelingLanguage Modellingspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
EmoTale: An Enacted Speech-emotion Dataset in Danish
While multiple emotional speech corpora exist for commonly spoken languages, there is a lack of functional datasets for smaller (spoken) languages, such as Danish. To our knowledge, Danish Emotional Speech (DES), publish…
Speech Emotion RecognitionInvestigating the Impact of ASR Errors on Spoken Implicit Discourse Relation Recognition
We present an empirical study investigating the influence of automatic speech recognition (ASR) errors on the spoken implicit discourse relation recognition (IDRR) task. We construct a spoken dataset for this task based …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Relationspeech-recognition+1