paper-with-me

홈 › Papers

Scheduled Dialog Policy Learning: An Automatic Curriculum Learning Framework for Task-oriented Dialog System

2021-08-01 · Findings (ACL) 2021 8 · Sihong Liu, Jinchao Zhang, Keqing He, Weiran Xu, Jie zhou
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Scheduled Curiosity-Deep Dyna-Q: Efficient Exploration for Dialog Policy Learning

2024-01-31 · Xuecheng Niu, Akinori Ito, Takashi Nose

Training task-oriented dialog agents based on reinforcement learning is time-consuming and requires a large number of interactions with real users. How to grasp dialog policy within limited dialog experiences remains an …

Efficient ExplorationModel-based Reinforcement LearningQ-Learningreinforcement-learning

Automatic Curriculum Learning With Over-repetition Penalty for Dialogue Policy Learning

2020-12-28 · Yangyang Zhao, Zhenyu Wang, Zhenhua Huang

Dialogue policy learning based on reinforcement learning is difficult to be applied to real users to train dialogue agents from scratch because of the high cost. User simulators, which choose random user goals for the di…

A Versatile Adaptive Curriculum Learning Framework for Task-oriented Dialogue Policy Learning

2022-07-01 · Findings (NAACL) 2022 7 · Yang Zhao, Hua Qin, Wang Zhenyu, Changxi Zhu 외

Training a deep reinforcement learning-based dialogue policy with brute-force random sampling is costly. A new training paradigm was proposed to improve learning performance and efficiency by combining curriculum learnin…

Deep Reinforcement Learning

Curriculum Learning Based on Reward Sparseness for Deep Reinforcement Learning of Task Completion Dialogue Management

2018-10-01 · WS 2018 10 · Atsushi Saito

Learning from sparse and delayed reward is a central issue in reinforcement learning. In this paper, to tackle reward sparseness problem of task oriented dialogue management, we propose a curriculum based approach on the…

Deep Reinforcement LearningDialogue ManagementInformation RetrievalManagement+5

TA&AT: Enhancing Task-Oriented Dialog with Turn-Level Auxiliary Tasks and Action-Tree Based Scheduled Sampling

2024-01-28 · Longxiang Liu, Xiuxing Li, Yang Feng

Task-oriented dialog systems have witnessed substantial progress due to conversational pre-training techniques. Yet, two significant challenges persist. First, most systems primarily utilize the latest turn's state label…

Decoder