paper-with-me

홈 › Papers

A Visually-grounded First-person Dialogue Dataset with Verbal and Non-verbal Responses

2020-11-01 · EMNLP 2020 11 · Hisashi Kamezawa, Noriki Nishida, Nobuyuki Shimizu, Takashi Miyazaki, Hideki Nakayama

In real-world dialogue, first-person visual information about where the other speakers are and what they are paying attention to is crucial to understand their intentions. Non-verbal responses also play an important role in social interactions. In this paper, we propose a visually-grounded first-person dialogue (VFD) dataset with verbal and non-verbal responses. The VFD dataset provides manually annotated (1) first-person images of agents, (2) utterances of human speakers, (3) eye-gaze locations of the speakers, and (4) the agents{'} verbal and non-verbal responses. We present experimental results obtained using the proposed VFD dataset and recent neural network models (e.g., BERT, ResNet). The results demonstrate that first-person vision helps neural network models correctly understand human intentions, and the production of non-verbal responses is a challenging task like that of verbal responses. Our dataset is publicly available.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Linguistic Analysis of Visually Grounded Dialogues Based on Spatial Expressions

2020-10-07 · Findings of the Association for Computational Linguistics 2020 · Takuma Udagawa, Takato Yamazaki, Akiko Aizawa

Recent models achieve promising results in visually grounded dialogues. However, existing datasets often contain undesirable biases and lack sophisticated linguistic analyses, which make it difficult to understand how we…

Coreference ResolutionNatural Language Visual GroundingSpatial Relation Recognition

VDialogUE: A Unified Evaluation Benchmark for Visually-grounded Dialogue

2023-09-14 · Yunshui Li, Binyuan Hui, Zhaochao Yin, Wanwei He 외

Visually-grounded dialog systems, which integrate multiple modes of communication such as text and visual inputs, have become an increasingly popular area of investigation. However, the absence of a standardized evaluati…

MPCHAT: Towards Multimodal Persona-Grounded Conversation

2023-05-27 · Jaewoo Ahn, Yeda Song, Sangdoo Yun, Gunhee Kim

In order to build self-consistent personalized dialogue agents, previous research has mostly focused on textual persona that delivers personal facts or personalities. However, to fully describe the multi-faceted nature o…

Speaker Identification

A Visually-Aware Conversational Robot Receptionist

2022-09-01 · SIGDIAL (ACL) 2022 9 · Nancie Gunson, Daniel Hernandez Garcia, Weronika Sieińska, Angus Addlesee 외

Socially Assistive Robots (SARs) have the potential to play an increasingly important role in a variety of contexts including healthcare, but most existing systems have very limited interactive capabilities. We will demo…

Question Answering

Transferable Persona-Grounded Dialogues via Grounded Minimal Edits

2021-09-16 · EMNLP 2021 11 · Chen Henry Wu, Yinhe Zheng, Xiaoxi Mao, Minlie Huang

Grounded dialogue models generate responses that are grounded on certain concepts. Limited by the distribution of grounded dialogue data, models trained on such data face the transferability challenges in terms of the da…