paper-with-me

홈 › Papers

Neural Joking Machine : Humorous image captioning

2018-05-30 · Kota Yoshida, Munetaka Minoguchi, Kenichiro Wani, Akio Nakamura, Hirokatsu Kataoka

What is an effective expression that draws laughter from human beings? In the present paper, in order to consider this question from an academic standpoint, we generate an image caption that draws a "laugh" by a computer. A system that outputs funny captions based on the image caption proposed in the computer vision field is constructed. Moreover, we also propose the Funny Score, which flexibly gives weights according to an evaluation database. The Funny Score more effectively brings out "laughter" to optimize a model. In addition, we build a self-collected BoketeDB, which contains a theme (image) and funny caption (text) posted on "Bokete", which is an image Ogiri website. In an experiment, we use BoketeDB to verify the effectiveness of the proposed method by comparing the results obtained using the proposed method and those obtained using MS COCO Pre-trained CNN+LSTM, which is the baseline and idiot created by humans. We refer to the proposed method, which uses the BoketeDB pre-trained model, as the Neural Joking Machine (NJM).

📄 PDF Abstract BibTeX arXiv:1805.11850

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioning

Similar Papers 제목 키워드 기반

OxfordTVG-HIC: Can Machine Make Humorous Captions from Images?

2023-07-21 · ICCV 2023 1 · Runjia Li, Shuyang Sun, Mohamed Elhoseiny, Philip Torr

This paper presents OxfordTVG-HIC (Humorous Image Captions), a large-scale dataset for humour generation and understanding. Humour is an abstract, subjective, and context-dependent cognitive construct involving several c…

DiversityImage Captioning

Culture-Aware Humorous Captioning: Multimodal Humor Generation across Cultural Contexts

2026-04-20 · Run Xu, Lu Li, Rongzhao Zhang, Jie Xu arxiv

Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over explicit cultural context, making it difficult to jointly maintain …

multimodal generation

"Factual" or "Emotional": Stylized Image Captioning with Adaptive Learning and Attention

2018-07-10 · Tianlang Chen, Zhongping Zhang, Quanzeng You, Chen Fang 외

Generating stylized captions for an image is an emerging topic in image captioning. Given an image as input, it requires the system to generate a caption that has a specific style (e.g., humorous, romantic, positive, and…

Image Captioning

``Factual'' or ``Emotional'': Stylized Image Captioning with Adaptive Learning and Attention

2018-09-01 · ECCV 2018 9 · Tianlang Chen, Zhongping Zhang, Quanzeng You, Chen Fang 외

Generating stylized captions for an image is an emerging topic in image captioning. Given an image as input, it requires the system to generate a caption that has a specific style (e.g., humorous, romantic, positive, and…

Image Captioning

Exploiting Image–Text Synergy for Contextual Image Captioning

2021-04-01 · EACL (LANTERN) 2021 4 · Sreyasi Nag Chowdhury, Rajarshi Bhowmik, Hareesh Ravi, Gerard de Melo 외

Modern web content - news articles, blog posts, educational resources, marketing brochures - is predominantly multimodal. A notable trait is the inclusion of media such as images placed at meaningful locations within a t…

ArticlesImage CaptioningMarketing