paper-with-me

홈 › Papers

Generative Visual Dialogue System via Adaptive Reasoning and Weighted Likelihood Estimation

2019-02-26 · Heming Zhang, Shalini Ghosh, Larry Heck, Stephen Walsh, Junting Zhang, Jie Zhang, C. -C. Jay Kuo

The key challenge of generative Visual Dialogue (VD) systems is to respond to human queries with informative answers in natural and contiguous conversation flow. Traditional Maximum Likelihood Estimation (MLE)-based methods only learn from positive responses but ignore the negative responses, and consequently tend to yield safe or generic responses. To address this issue, we propose a novel training scheme in conjunction with weighted likelihood estimation (WLE) method. Furthermore, an adaptive multi-modal reasoning module is designed, to accommodate various dialogue scenarios automatically and select relevant information accordingly. The experimental results on the VisDial benchmark demonstrate the superiority of our proposed algorithm over other state-of-the-art approaches, with an improvement of 5.81% on recall@10.

📄 PDF Abstract BibTeX arXiv:1902.09818

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Dialog

Similar Papers 제목 키워드 기반

Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance

2025-07-14 · Kyungtae Han, Yitao Chen, Rohit Gupta, Onur Altintas

While autonomous driving technologies continue to advance, current Advanced Driver Assistance Systems (ADAS) remain limited in their ability to interpret scene context or engage with drivers through natural language. The…

Autonomous Driving

KBGN: Knowledge-Bridge Graph Network for Adaptive Vision-Text Reasoning in Visual Dialogue

2020-08-11 · Xiaoze Jiang, Siyi Du, Zengchang Qin, Yajing Sun 외

Visual dialogue is a challenging task that needs to extract implicit information from both visual (image) and textual (dialogue history) contexts. Classical approaches pay more attention to the integration of the current…

Information RetrievalRetrieval

EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Spoken Dialogue Systems

2025-08-25 · Jingwen Liu, Kan Jen Cheng, Jiachen Lian, Akshay Anand 외 arxiv

Speech emotions play a crucial role in human-computer interaction, shaping engagement and context-aware communication. Despite recent advances in spoken dialogue systems, a holistic system for evaluating emotional reason…

DVD: A Diagnostic Dataset for Multi-step Reasoning in Video Grounded Dialogue

2021-01-01 · ACL 2021 5 · Hung Le, Chinnadhurai Sankar, Seungwhan Moon, Ahmad Beirami 외

A video-grounded dialogue system is required to understand both dialogue, which contains semantic dependencies from turn to turn, and video, which contains visual cues of spatial and temporal scene variations. Building s…

DiagnosticObject TrackingVisual Reasoning

Are You Talking to Me? Reasoned Visual Dialog Generation through Adversarial Learning

2017-11-21 · CVPR 2018 6 · Qi Wu, Peng Wang, Chunhua Shen, Ian Reid 외

The Visual Dialogue task requires an agent to engage in a conversation about an image with a human. It represents an extension of the Visual Question Answering task in that the agent needs to answer a question about an i…

Question AnsweringReinforcement LearningVisual DialogVisual Question Answering+1