paper-with-me

Scene-Aware Dialogue

1개 벤치마크 · 논문 8편 · 이 태스크의 논문 보기 →

Benchmarks

AVSD

결과 1개

Most implemented

Audio-Visual Scene-Aware Dialog

2019-01-25 · 구현 2개

An Embodied Generalist Agent in 3D World

2023-11-18 · 구현 1개

Papers

An Embodied Generalist Agent in 3D World

2023-11-18 · Jiangyong Huang, Silong Yong, Xiaojian Ma, Xiongkun Linghu 외

Leveraging massive knowledge from large language models (LLMs), recent machine learning models show notable successes in general-purpose task solving in diverse domains such as computer vision and robotics. However, seve…

3D dense captioning3D Question Answering (3D-QA)Question AnsweringRobot Manipulation+3

Maintaining Common Ground in Dynamic Environments

2021-05-29 · Takuma Udagawa, Akiko Aizawa

Common grounding is the process of creating and maintaining mutual understandings, which is a critical aspect of sophisticated human communication. While various task settings have been proposed in existing literature, t…

End-To-End Dialogue ModellingGoal-Oriented Dialogue SystemsScene-Aware Dialogue

Multimodal Dialogue State Tracking By QA Approach with Data Augmentation

2020-07-20 · Xiangyang Mou, Brandyn Sigouin, Ian Steenstra, Hui Su

Recently, a more challenging state tracking task, Audio-Video Scene-Aware Dialogue (AVSD), is catching an increasing amount of attention among researchers. Different from purely text-based dialogue state tracking, the di…

Data AugmentationDecoderDialogue State TrackingOpen-Domain Question Answering+2

Multi-step Joint-Modality Attention Network for Scene-Aware Dialogue System

2020-01-17 · Yun-Wei Chu, Kuan-Yen Lin, Chao-Chun Hsu, Lun-Wei Ku

Understanding dynamic scenes and dialogue contexts in order to converse with users has been challenging for multimodal dialogue systems. The 8-th Dialog System Technology Challenge (DSTC8) proposed an Audio Visual Scene-…

Scene-Aware Dialogue

Entropy-Enhanced Multimodal Attention Model for Scene-Aware Dialogue Generation

2019-08-22 · Kuan-Yen Lin, Chao-Chun Hsu, Yun-Nung Chen, Lun-Wei Ku

With increasing information from social media, there are more and more videos available. Therefore, the ability to reason on a video is important and deserves to be discussed. TheDialog System Technology Challenge (DSTC7…

Dialogue GenerationScene-Aware Dialogue

Reactive Multi-Stage Feature Fusion for Multimodal Dialogue Modeling

2019-08-14 · Yi-Ting Yeh, Tzu-Chuan Lin, Hsiao-Hua Cheng, Yu-Hsuan Deng 외

Visual question answering and visual dialogue tasks have been increasingly studied in the multimodal field towards more practical real-world scenarios. A more challenging task, audio visual scene-aware dialogue (AVSD), i…

Question AnsweringScene-Aware DialogueVisual DialogVisual Question Answering+1

전체 8편 보기 →