Towards End-to-End Learning for Dialog State Tracking and Management using Deep Reinforcement Learning
This paper presents an end-to-end framework for task-oriented dialog systems using a variant of Deep Recurrent Q-Networks (DRQN). The model is able to interface with a relational database and jointly learn policies for both language understanding and dialog strategy. Moreover, we propose a hybrid algorithm that combines the strength of reinforcement learning and supervised learning to achieve faster learning speed. We evaluated the proposed model on a 20 Question Game conversational game simulator. Results show that the proposed method outperforms the modular-based baseline and learns a distributed representation of the latent dialog state.
Code (1)
Tasks
Deep Reinforcement Learningdialog state trackingManagementreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Distributed Structured Actor-Critic Reinforcement Learning for Universal Dialogue Management
The task-oriented spoken dialogue system (SDS) aims to assist a human user in accomplishing a specific task (e.g., hotel booking). The dialogue management is a core part of SDS. There are two main missions in dialogue ma…
Decision MakingDeep Reinforcement LearningDialogue ManagementManagement+3Show Us the Way: Learning to Manage Dialog from Demonstrations
We present our submission to the End-to-End Multi-Domain Dialog Challenge Track of the Eighth Dialog System Technology Challenge. Our proposed dialog system adopts a pipeline architecture, with distinct components for Na…
dialog state trackingManagementNatural Language UnderstandingQ-Learning+4Deep Reinforcement Learning for On-line Dialogue State Tracking
Dialogue state tracking (DST) is a crucial module in dialogue management. It is usually cast as a supervised training problem, which is not convenient for on-line optimization. In this paper, a novel companion teaching b…
Deep Reinforcement LearningDialogue ManagementDialogue State TrackingManagement+4Spectral decomposition method of dialog state tracking via collective matrix factorization
The task of dialog management is commonly decomposed into two sequential subtasks: dialog state tracking and dialog policy learning. In an end-to-end dialog system, the aim of dialog state tracking is to accurately estim…
dialog state trackingManagementNatural Language Understandingspeech-recognition+1Empathetic Response Generation with State Management
A good empathetic dialogue system should first track and understand a user's emotion and then reply with an appropriate emotion. However, current approaches to this task either focus on improving the understanding of use…
Dialogue ManagementEmpathetic Response GenerationManagementResponse Generation+1