A dataset for resolving referring expressions in spoken dialogue via contextual query rewrites (CQR)
We present Contextual Query Rewrite (CQR) a dataset for multi-domain task-oriented spoken dialogue systems that is an extension of the Stanford dialog corpus (Eric et al., 2017a). While previous approaches have addressed the issue of diverse schemas by learning candidate transformations (Naik et al., 2018), we instead model the reference resolution task as a user query reformulation task, where the dialog state is serialized into a natural language query that can be executed by the downstream spoken language understanding system. In this paper, we describe our methodology for creating the query reformulation extension to the dialog corpus, and present an initial set of experiments to establish a baseline for the CQR task. We have released the corpus to the public [1] to support further research in this area.
Code (1)
Tasks
Spoken Dialogue SystemsSpoken Language UnderstandingSimilar Papers 제목 키워드 기반
Scaling Multi-Domain Dialogue State Tracking via Query Reformulation
We present a novel approach to dialogue state tracking and referring expression resolution tasks. Successful contextual understanding of multi-turn spoken dialogues requires resolving referring expressions across turns a…
Dialogue State TrackingMulti-domain Dialogue State TrackingMulti-Task LearningReferring Expression+1Refer-iTTS: A System for Referring in Spoken Installments to Objects in Real-World Images
Current referring expression generation systems mostly deliver their output as one-shot, written expressions. We present on-going work on incremental generation of spoken expressions referring to objects in real-world im…
Referring ExpressionReferring expression generationSpeech SynthesisText Generation+3SpaceRefNet: a neural approach to spatial reference resolution in a real city environment
Adding interactive capabilities to pedestrian wayfinding systems in the form of spoken dialogue will make them more natural to humans. Such an interactive wayfinding system needs to continuously understand and interpret …
PentoRef: A Corpus of Spoken References in Task-oriented Dialogues
PentoRef is a corpus of task-oriented dialogues collected in systematically manipulated settings. The corpus is multilingual, with English and German sections, and overall comprises more than 20000 utterances. The dialog…
Resolving Referring Expressions in Images With Labeled Elements
Images may have elements containing text and a bounding box associated with them, for example, text identified via optical character recognition on a computer screen image, or a natural image with labeled objects. We pre…
Optical Character RecognitionOptical Character Recognition (OCR)Referring Expression