paper-with-me

Papers

Improving Device Directedness Classification of Utterances with Semantic Lexical Features

2020-09-29 · Kellen Gillespie, Ioannis C. Konstantakopoulos, Xingzhi Guo, Vishal Thanvantri Vasudevan, Abhinav Sethy

User interactions with personal assistants like Alexa, Google Home and Siri are typically initiated by a wake term or wakeword. Several personal assistants feature "follow-up" modes that allow users to make additional interactions without the need of a wakeword. For the system to only respond when appropriate, and to ignore speech not intended for it, utterances must be classified as device-directed or non-device-directed. State-of-the-art systems have largely used acoustic features for this task, while others have used only lexical features or have added LM-based lexical features. We propose a directedness classifier that combines semantic lexical features with a lightweight acoustic feature and show it is effective in classifying directedness. The mixed-domain lexical and acoustic feature model is able to achieve 14% relative reduction of EER over a state-of-the-art acoustic-only baseline model. Finally, we successfully apply transfer learning and semi-supervised learning to the model to improve accuracy even further.

📄 PDF Abstract BibTeX arXiv:2010.01949

Code (0)

등록된 구현이 없습니다.

Tasks

General ClassificationTransfer Learning

Similar Papers 제목 키워드 기반

Device Directedness with Contextual Cues for Spoken Dialog Systems

2022-11-23 · Dhanush Bekal, Sundararajan Srinivasan, Sravan Bodapati, Srikanth Ronanki 외

In this work, we define barge-in verification as a supervised learning task where audio-only information is used to classify user spoken dialogue into true and false barge-ins. Following the success of pre-trained models…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Representation Learningspeech-recognition+1

Machine Semiotics

2020-08-24 · Peter beim Graben, Markus Huber-Liebl, Peter Klimczak, Günther Wirsching

Recognizing a basic difference between the semiotics of humans and machines presents a possibility to overcome the shortcomings of current speech assistive devices. For the machine, the meaning of a (human) utterance is …

Implicaturesspeech-recognitionSpeech Recognition

Measuring Goal-Directedness

2024-12-06 · Matt MacDermott, James Fox, Francesco Belardinelli, Tom Everitt

We define maximum entropy goal-directedness (MEG), a formal measure of goal-directedness in causal models and Markov decision processes, and give algorithms for computing it. Measuring goal-directedness is important, as …

Delexicalized Paraphrase Generation

2020-12-04 · COLING 2020 8 · Boya Yu, Konstantine Arkoudas, Wael Hamza

We present a neural model for paraphrasing and train it to generate delexicalized sentences. We achieve this by creating training data in which each input is paired with a number of reference paraphrases. These sets of r…

Data Augmentationintent-classificationIntent Classificationnamed-entity-recognition+4

How Would You Say It? Eliciting Lexically Diverse Dialogue for Supervised Semantic Parsing

2017-08-01 · WS 2017 8 · Ravich, Abhilasha er, Thomas Manzini, Matthias Grabmair 외

Building dialogue interfaces for real-world scenarios often entails training semantic parsers starting from zero examples. How can we build datasets that better capture the variety of ways users might phrase their querie…

Semantic Parsing