paper-with-me

홈 › Papers

On-line Dialogue Policy Learning with Companion Teaching

2017-04-01 · EACL 2017 4 · Lu Chen, Runzhe Yang, Cheng Chang, Zihao Ye, Xiang Zhou, Kai Yu

On-line dialogue policy learning is the key for building evolvable conversational agent in real world scenarios. Poor initial policy can easily lead to bad user experience and consequently fail to attract sufficient users for policy training. A novel framework, companion teaching, is proposed to include a human teacher in the dialogue policy training loop to address the cold start problem. Here, dialogue policy is trained using not only user{'}s reward, but also teacher{'}s example action as well as estimated immediate reward at turn level. Simulation experiments showed that, with small number of human teaching dialogues, the proposed approach can effectively improve user experience at the beginning and smoothly lead to good performance with more user interaction data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue Management

Similar Papers 제목 키워드 기반

Affordable On-line Dialogue Policy Learning

2017-09-01 · EMNLP 2017 9 · Cheng Chang, Runzhe Yang, Lu Chen, Xiang Zhou 외

The key to building an evolvable dialogue system in real-world scenarios is to ensure an affordable on-line dialogue policy learning, which requires the on-line learning process to be safe, efficient and economical. But …

Dialogue Management

Deep Reinforcement Learning for On-line Dialogue State Tracking

2020-09-22 · Zhi Chen, Lu Chen, Xiang Zhou, Kai Yu

Dialogue state tracking (DST) is a crucial module in dialogue management. It is usually cast as a supervised training problem, which is not convenient for on-line optimization. In this paper, a novel companion teaching b…

Deep Reinforcement LearningDialogue ManagementDialogue State TrackingManagement+4

Agent-Aware Dropout DQN for Safe and Efficient On-line Dialogue Policy Learning

2017-09-01 · EMNLP 2017 9 · Lu Chen, Xiang Zhou, Cheng Chang, Runzhe Yang 외

Hand-crafted rules and reinforcement learning (RL) are two popular choices to obtain dialogue policy. The rule-based policy is often reliable within predefined scope but not self-adaptable, whereas RL is evolvable with d…

Automatic Speech Recognition (ASR)Dialogue ManagementReinforcement LearningReinforcement Learning (RL)+3

An AI-based Learning Companion Promoting Lifelong Learning Opportunities for All

2021-11-16 · Maria Perez-Ortiz, Erik Novak, Sahan Bulathwela, John Shawe-Taylor

Artifical Intelligence (AI) in Education has great potential for building more personalised curricula, as well as democratising education worldwide and creating a Renaissance of new ways of teaching and learning. We beli…

AllLifelong learning

MoodBench 1.0: An Evaluation Benchmark for Emotional Companionship Dialogue Systems

2025-11-24 · Haifeng Jing, Yujie Hou, Junfei Liu, Rui Xie 외 arxiv

With the rapid development of Large Language Models, dialogue systems are shifting from information tools to emotional companions, heralding the era of Emotional Companionship Dialogue Systems (ECDs) that provide persona…