paper-with-me

홈 › Papers

Mind Your Language: Learning Visually Grounded Dialog in a Multi-Agent Setting

2018-05-25 · Anonymous

The task of visually grounded dialog involves learning goal-oriented cooperative dialog between autonomous agents who exchange information about a scene through several rounds of questions and answers. We posit that requiring agents to adhere to rules of human language while also maximizing information exchange is an ill-posed problem, and observe that humans do not stray from a common language, because they are social creatures and have to communicate with many people everyday, and it is far easier to stick to a common language even at the cost of some efficiency loss. Using this as inspiration, we propose and evaluate a multi-agent dialog framework where each agent interacts with, and learns from, multiple agents, and show that this results in more relevant and coherent dialog (as judged by human evaluators) without sacrificing task performance (as judged by quantitative metrics).

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Speaking the Language of Your Listener: Audience-Aware Adaptation via Plug-and-Play Theory of Mind

2023-05-31 · Ece Takmaz, Nicolo' Brandizzi, Mario Giulianelli, Sandro Pezzelle 외

Dialogue participants may have varying levels of knowledge about the topic under discussion. In such cases, it is essential for speakers to adapt their utterances by taking their audience into account. Yet, it is an open…

Language ModelingLanguage ModellingOpen-Ended Question AnsweringText Generation

Dialog without Dialog Data: Learning Visual Dialog Agents from VQA Data

2020-07-24 · NeurIPS 2020 12 · Michael Cogswell, Jiasen Lu, Rishabh Jain, Stefan Lee 외

Can we develop visually grounded dialog agents that can efficiently adapt to new tasks without forgetting how to talk to people? Such agents could leverage a larger variety of existing data to generalize to new tasks, mi…

Visual DialogVisual Question Answering (VQA)

Resolving References in Visually-Grounded Dialogue via Text Generation

2023-09-23 · SIGdial 2023 9 · Bram Willemsen, Livia Qian, Gabriel Skantze

Vision-language models (VLMs) have shown to be effective at image retrieval based on simple text queries, but text-image retrieval based on conversational input remains a challenge. Consequently, if we want to use VLMs f…

Image RetrievalLanguage ModelingLanguage ModellingLarge Language Model+2

Affective Visual Dialog: A Large-Scale Benchmark for Emotional Reasoning Based on Visually Grounded Conversations

2023-08-30 · Kilichbek Haydarov, Xiaoqian Shen, Avinash Madasu, Mahmoud Salem 외

We introduce Affective Visual Dialog, an emotion explanation and reasoning task as a testbed for research on understanding the formation of emotions in visually grounded conversations. The task involves three skills: (1)…

Explanation GenerationQuestion AnsweringVisual Dialog

Incremental Generation of Visually Grounded Language in Situated Dialogue (demonstration system)

2016-09-01 · WS 2016 9 · Yanchao Yu, Arash Eshghi, Oliver Lemon
Text Generation