paper-with-me

Papers

The JDDC 2.0 Corpus: A Large-Scale Multimodal Multi-Turn Chinese Dialogue Dataset for E-commerce Customer Service

2021-09-27 · Nan Zhao, Haoran Li, Youzheng Wu, Xiaodong He, BoWen Zhou

With the development of the Internet, more and more people get accustomed to online shopping. When communicating with customer service, users may express their requirements by means of text, images, and videos, which precipitates the need for understanding these multimodal information for automatic customer service systems. Images usually act as discriminators for product models, or indicators of product failures, which play important roles in the E-commerce scenario. On the other hand, detailed information provided by the images is limited, and typically, customer service systems cannot understand the intents of users without the input text. Thus, bridging the gap of the image and text is crucial for the multimodal dialogue task. To handle this problem, we construct JDDC 2.0, a large-scale multimodal multi-turn dialogue dataset collected from a mainstream Chinese E-commerce platform (JD.com), containing about 246 thousand dialogue sessions, 3 million utterances, and 507 thousand images, along with product knowledge bases and image category annotations. We present the solutions of top-5 teams participating in the JDDC multimodal dialogue challenge based on this dataset, which provides valuable insights for further researches on the multimodal dialogue task.

📄 PDF Abstract BibTeX arXiv:2109.12913

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

The JDDC Corpus: A Large-Scale Multi-Turn Chinese Dialogue Dataset for E-commerce Customer Service

2019-11-22 · LREC 2020 5 · Meng Chen, Ruixue Liu, Lei Shen, Shaozu Yuan 외

Human conversations are complicated and building a human-like dialogue agent is an extremely challenging task. With the rapid development of deep learning techniques, data-driven models become more and more prevalent whi…

Question AnsweringRetrieval

OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text

2024-06-12 · Qingyun Li, Zhe Chen, Weiyun Wang, Wenhai Wang 외

Image-text interleaved data, consisting of multiple images and texts arranged in a natural document format, aligns with the presentation paradigm of internet data and closely resembles human reading habits. Recent studie…

In-Context Learning

A Large Scale Speech Sentiment Corpus

2020-05-01 · LREC 2020 5 · Eric Chen, Zhiyun Lu, Hao Xu, Liangliang Cao 외

We present a multimodal corpus for sentiment analysis based on the existing Switchboard-1 Telephone Speech Corpus released by the Linguistic Data Consortium. This corpus extends the Switchboard-1 Telephone Speech Corpus …

Sentiment Analysis

EgMM-Corpus: A Multimodal Vision-Language Dataset for Egyptian Culture

2025-10-17 · Mohamed Gamil, Abdelrahman Elsayed, Abdelrahman Lila, Ahmed Gad 외 arxiv

Despite recent advances in AI, multimodal culturally diverse datasets are still limited, particularly for regions in the Middle East and Africa. In this paper, we introduce EgMM-Corpus, a multimodal dataset dedicated to …

Multimodal Knowledge Learning for Named Entity Disambiguation

2021-08-17 · ACL ARR August 2021 8 · Anonymous

With the popularity of online social medias in recent years, massive-scale multimodal information has brought new challenges to traditional Named Entity Disambiguation (NED) tasks. Recently, Multimodal Named Entity Disam…

Entity DisambiguationMeta-LearningTransfer Learning