Scheduled Dialog Policy Learning: An Automatic Curriculum Learning Framework for Task-oriented Dialog System
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Scheduled Curiosity-Deep Dyna-Q: Efficient Exploration for Dialog Policy Learning
Training task-oriented dialog agents based on reinforcement learning is time-consuming and requires a large number of interactions with real users. How to grasp dialog policy within limited dialog experiences remains an …
Efficient ExplorationModel-based Reinforcement LearningQ-Learningreinforcement-learningAutomatic Curriculum Learning With Over-repetition Penalty for Dialogue Policy Learning
Dialogue policy learning based on reinforcement learning is difficult to be applied to real users to train dialogue agents from scratch because of the high cost. User simulators, which choose random user goals for the di…
A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning
Training a deep reinforcement learning-based dialogue policy with brute-force random sampling is costly. A new training paradigm was proposed to improve learning performance and efficiency by combining curriculum learnin…
Deep Reinforcement LearningCurriculum Learning Based on Reward Sparseness for Deep Reinforcement Learning of Task Completion Dialogue Management
Learning from sparse and delayed reward is a central issue in reinforcement learning. In this paper, to tackle reward sparseness problem of task oriented dialogue management, we propose a curriculum based approach on the…
Deep Reinforcement LearningDialogue ManagementInformation RetrievalManagement+5TA&AT: Enhancing Task-Oriented Dialog with Turn-Level Auxiliary Tasks and Action-Tree Based Scheduled Sampling
Task-oriented dialog systems have witnessed substantial progress due to conversational pre-training techniques. Yet, two significant challenges persist. First, most systems primarily utilize the latest turn's state label…
Decoder