Investigating Reinforcement Learning for Communication Strategies in a Task-Initiative Setting
Many conversational domains require the system to present nuanced information to users. Such systems must follow up what they say to address clarification questions and repair misunderstandings. In this work, we explore this interactive strategy in a referential communication task. Using simulation, we analyze the communication trade-offs between initial presentation and subsequent followup as a function of user clarification strategy, and compare the performance of several baseline strategies to policies derived by reinforcement learning. We find surprising advantages to coherence-based representations of dialogue strategy, which bring minimal data requirements, explainable choices, and strong audit capabilities, but incur little loss in predicted outcomes across a wide range of user models.
Code (0)
등록된 구현이 없습니다.
Tasks
reinforcement-learningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Modeling Non-Cooperative Dialogue: Theoretical and Empirical Insights
Investigating cooperativity of interlocutors is central in studying pragmatics of dialogue. Models of conversation that only assume cooperative agents fail to explain the dynamics of strategic conversations. Thus, we inv…
Learning TheoryCreative Wand: A System to Study Effects of Communications in Co-Creative Settings
Recent neural generation systems have demonstrated the potential for procedurally generating game content, images, stories, and more. However, most neural generation algorithms are "uncontrolled" in the sense that the us…
From SPMRL to NMRL: What Did We Learn (and Unlearn) in a Decade of Parsing Morphologically-Rich Languages (MRLs)?
It has been exactly a decade since the first establishment of SPMRL, a research initiative unifying multiple research efforts to address the peculiar challenges of Statistical Parsing for Morphologically-Rich Languages (…
Effective Social Chatbot Strategies for Increasing User Initiative
Many existing chatbots do not effectively support mixed initiative, forcing their users to either respond passively or lead constantly. We seek to improve this experience by introducing new mechanisms to encourage user i…
ChatbotDiversityVehicle Communication Strategies for Simulated Highway Driving
Interest in emergent communication has recently surged in Machine Learning. The focus of this interest has largely been either on investigating the properties of the learned protocol or on utilizing emergent communicatio…
BIG-bench Machine LearningSelf-Driving Cars