paper-with-me

홈 › Papers

ChatPainter: Improving Text to Image Generation using Dialogue

2018-02-22 · Shikhar Sharma, Dendi Suhubdy, Vincent Michalski, Samira Ebrahimi Kahou, Yoshua Bengio

Synthesizing realistic images from text descriptions on a dataset like Microsoft Common Objects in Context (MS COCO), where each image can contain several objects, is a challenging task. Prior work has used text captions to generate images. However, captions might not be informative enough to capture the entire image and insufficient for the model to be able to understand which objects in the images correspond to which words in the captions. We show that adding a dialogue that further describes the scene leads to significant improvement in the inception score and in the quality of generated images on the MS COCO dataset.

📄 PDF Abstract BibTeX arXiv:1802.08216

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationText to Image GenerationText-to-Image Generation

Similar Papers 제목 키워드 기반

Open Domain Dialogue Generation with Latent Images

2020-04-04 · Ze Yang, Wei Wu, Huang Hu, Can Xu 외

We consider grounding open domain dialogues with images. Existing work assumes that both an image and a textual context are available, but image-grounded dialogues by nature are more difficult to obtain than textual dial…

Dialogue GenerationImage GenerationResponse GenerationText to Image Generation+1

Multimodal Dialogue Response Generation

2021-10-16 · ACL 2022 5 · Qingfeng Sun, Yujing Wang, Can Xu, Kai Zheng 외

Responsing with image has been recognized as an important capability for an intelligent conversational agent. Yet existing works only focus on exploring the multimodal dialogue models which depend on retrieval-based meth…

Dialogue GenerationResponse GenerationRetrieval

An End-to-End Model for Photo-Sharing Multi-modal Dialogue Generation

2024-08-16 · Peiming Guo, Sinuo Liu, Yanzhao Zhang, Dingkun Long 외

Photo-Sharing Multi-modal dialogue generation requires a dialogue agent not only to generate text responses but also to share photos at the proper moment. Using image text caption as the bridge, a pipeline model integrat…

Dialogue GenerationImage GenerationLanguage ModelingLanguage Modelling+3

BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation

2024-08-12 · Hee Suk Yoon, Eunseop Yoon, Joshua Tian Jin Tee, Kang Zhang 외

Multimodal Dialogue Response Generation (MDRG) is a recently proposed task where the model needs to generate responses in texts, images, or a blend of both based on the dialogue context. Due to the lack of a large-scale …

Response Generation

Constructing Multi-Modal Dialogue Dataset by Replacing Text with Semantically Relevant Images

2021-07-19 · ACL 2021 5 · Nyoungwoo Lee, Suwon Shin, Jaegul Choo, Ho-Jin Choi 외

In multi-modal dialogue systems, it is important to allow the use of images as part of a multi-turn conversation. Training such dialogue systems generally requires a large-scale dataset consisting of multi-turn dialogues…

RetrievalSentence