paper-with-me

홈 › Papers

Region under Discussion for visual dialog

2021-11-01 · EMNLP 2021 11 · Mauricio Mazuecos, Franco M. Luque, Jorge Sánchez, Hernán Maina, Thomas Vadora, Luciana Benotti

Visual Dialog is assumed to require the dialog history to generate correct responses during a dialog. However, it is not clear from previous work how dialog history is needed for visual dialog. In this paper we define what it means for a visual question to require dialog history and we release a subset of the Guesswhat?! questions for which their dialog history completely changes their responses. We propose a novel interpretable representation that visually grounds dialog history: the Region under Discussion. It constrains the image’s spatial features according to a semantic representation of the history inspired by the information structure notion of Question under Discussion.We evaluate the architecture on task-specific multimodal models and the visual transformer model LXMERT.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Dialog

Methods 이 논문이 사용한 방법론

LXMERT LXMERT is a model for learning vision-and-language cross-modality representations. It consists of a Transformer model that consists three encoders: object relationship encoder, a…

Similar Papers 제목 키워드 기반

ICCV23 Visual-Dialog Emotion Explanation Challenge: SEU_309 Team Technical Report

2024-07-13 · Yixiao Yuan, Yingzhe Peng

The Visual-Dialog Based Emotion Explanation Generation Challenge focuses on generating emotion explanations through visual-dialog interactions in art discussions. Our approach combines state-of-the-art multi-modal models…

Explanation GenerationLanguage ModelingLanguage ModellingVisual Dialog

Can AI agents understand spoken conversations about data visualizations in online meetings?

2025-09-30 · Rizul Sharma, Tianyu Jiang, Seokki Lee, Jillian Aurisano arxiv

In this short paper, we present work evaluating an AI agent's understanding of spoken conversations about data visualizations in an online meeting scenario. There is growing interest in the development of AI-assistants t…

MoralDial: A Framework to Train and Evaluate Moral Dialogue Systems via Moral Discussions

2022-12-21 · Hao Sun, Zhexin Zhang, Fei Mi, Yasheng Wang 외

Morality in dialogue systems has raised great attention in research recently. A moral dialogue system aligned with users' values could enhance conversation engagement and user connections. In this paper, we propose a fra…

Annotating anaphoric phenomena in situated dialogue

2021-06-01 · ACL (mmsr, IWCS) 2021 6 · Sharid Loáiciga, Simon Dobnik, David Schlangen

In recent years several corpora have been developed for vision and language tasks. With this paper, we intend to start a discussion on the annotation of referential phenomena in situated dialogue. We argue that there is …

coreference-resolutionCoreference Resolution

DualVD: An Adaptive Dual Encoding Model for Deep Visual Understanding in Visual Dialogue

2019-11-17 · Xiaoze Jiang, Jing Yu, Zengchang Qin, Yingying Zhuang 외

Different from Visual Question Answering task that requires to answer only one question about an image, Visual Dialogue involves multiple questions which cover a broad range of visual content that could be related to any…

feature selectionQuestion AnsweringVisual DialogVisual Question Answering+1