Transferable Dialogue Systems and User Simulators
One of the difficulties in training dialogue systems is the lack of training data. We explore the possibility of creating dialogue data through the interaction between a dialogue system and a user simulator. Our goal is to develop a modelling framework that can incorporate new dialogue scenarios through self-play between the two agents. In this framework, we first pre-train the two agents on a collection of source domain dialogues, which equips the agents to converse with each other via natural language. With further fine-tuning on a small amount of target domain data, the agents continue to interact with the aim of improving their behaviors using reinforcement learning with structured reward functions. In experiments on the MultiWOZ dataset, two practical transfer learning problems are investigated: 1) domain adaptation and 2) single-to-multiple domain transfer. We demonstrate that the proposed framework is highly effective in bootstrapping the performance of the two agents in transfer learning. We also show that our method leads to improvements in dialogue system performance on complete datasets.
Code (1)
Tasks
Domain AdaptationTransfer LearningSimilar Papers 제목 키워드 기반
Generating Diverse Personas for User Simulators to Test Interview Dialogue Systems
This paper addresses the issue of the significant labor required to test interview dialogue systems. While interview dialogue systems are expected to be useful in various scenarios, like other dialogue systems, testing t…
Task-Oriented Dialogue SystemsMetaphorical User Simulators for Evaluating Task-oriented Dialogue Systems
Task-oriented dialogue systems (TDSs) are assessed mainly in an offline setting or through human evaluation. The evaluation is often limited to single-turn or is very time-intensive. As an alternative, user simulators th…
Conversational RecommendationTask-Oriented Dialogue SystemsMUST: A Framework for Training Task-oriented Dialogue Systems with Multiple User SimulaTors
Recent works try to optimize a Task-oriented Dialogue System with reinforcement learning (RL) by building user simulators. However, most of them only focus on training the dialogue system using a single user simulator. I…
Reinforcement Learning (RL)Task-Oriented Dialogue SystemsNeural User Simulation for Corpus-based Policy Optimisation of Spoken Dialogue Systems
User Simulators are one of the major tools that enable offline training of task-oriented dialogue systems. For this task the Agenda-Based User Simulator (ABUS) is often used. The ABUS is based on hand-crafted rules and i…
Dialogue ManagementDiversityReinforcement LearningSpoken Dialogue Systems+2Neural User Simulation for Corpus-based Policy Optimisation for Spoken Dialogue Systems
User Simulators are one of the major tools that enable offline training of task-oriented dialogue systems. For this task the Agenda-Based User Simulator (ABUS) is often used. The ABUS is based on hand-crafted rules and i…
DiversityReinforcement LearningSpoken Dialogue SystemsTask-Oriented Dialogue Systems+1