paper-with-me

Papers

Refine and Imitate: Reducing Repetition and Inconsistency in Persuasion Dialogues via Reinforcement Learning and Human Demonstration

2020-12-31 · Findings (EMNLP) 2021 11 · Weiyan Shi, Yu Li, Saurav Sahay, Zhou Yu

Persuasion dialogue systems reflect the machine's ability to make strategic moves beyond verbal communication, and therefore differentiate themselves from task-oriented or open-domain dialogue systems and have their own unique values. However, the repetition and inconsistency problems still persist in dialogue response generation and could substantially impact user experience and impede the persuasion outcome. Besides, although reinforcement learning (RL) approaches have achieved big success in strategic tasks such as games, they require a sophisticated user simulator to provide real-time feedback to the dialogue system, which limits the application of RL on persuasion dialogues. To address these issues towards a better persuasion dialogue system, we apply RL to refine a language model baseline without user simulators, and distill sentence-level information about repetition, inconsistency, and task relevance through rewards. Moreover, to better accomplish the persuasion task, the model learns from human demonstration to imitate human persuasion behavior and selects the most persuasive responses. Experiments show that our model outperforms previous state-of-the-art dialogue models on both automatic metrics and human evaluation results on a donation persuasion task, and generates more diverse, consistent and persuasive conversations according to the user feedback.

📄 PDF Abstract BibTeX arXiv:2012.15375

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModellingReinforcement Learning (RL)Response GenerationSentence

Similar Papers 제목 키워드 기반

Refine and Imitate: Reducing Repetition and Inconsistency in Dialogue Generation via Reinforcement Learning and Human Demonstration

2021-01-01 · Weiyan Shi, Yu Li, Saurav Sahay, Zhou Yu

Despite the recent success of large-scale language models on various downstream NLP tasks, the repetition and inconsistency problems still persist in dialogue response generation. Previous approaches have attempted to av…

Dialogue GenerationLanguage ModelingLanguage ModellingResponse Generation+1

Reasoning or Rambling? Exploring the Effect of Thinking on Agent Persuasion

2025-09-25 · Haodong Zhao, Jidong Li, Zhaomin Wu, Tianjie Ju 외 arxiv

Understanding persuasion is critical for the safety and reliability of multi-agent systems built on large language models (LLMs). This paper studies persuasion dynamics by contrasting general LLMs with Large Reasoning Mo…

RebuttalAgent: Strategic Persuasion in Academic Rebuttal via Theory of Mind

2026-01-22 · Zhitao He, Zongwei Lyu, Yi R Fung arxiv

Although artificial intelligence (AI) has become deeply integrated into various stages of the research workflow and achieved remarkable advancements, academic rebuttal remains a significant and underexplored challenge. T…

Reinforcement Learning

IVAC-P2L: Leveraging Irregular Repetition Priors for Improving Video Action Counting

2024-03-18 · Hang Wang, Zhi-Qi Cheng, Youtian Du, Lei Zhang

Video Action Counting (VAC) is crucial in analyzing sports, fitness, and everyday activities by quantifying repetitive actions in videos. However, traditional VAC methods have overlooked the complexity of action repetiti…

Mitigating Manipulation and Enhancing Persuasion: A Reflective Multi-Agent Approach for Legal Argument Generation

2025-06-03 · Li Zhang, Kevin D. Ashley

Large Language Models (LLMs) are increasingly explored for legal argument generation, yet they pose significant risks of manipulation through hallucination and ungrounded persuasion, and often fail to utilize provided fa…

Hallucination