paper-with-me

홈 › Papers

What do you MEME? Generating Explanations for Visual Semantic Role Labelling in Memes

2022-12-01 · Shivam Sharma, Siddhant Agarwal, Tharun Suresh, Preslav Nakov, Md. Shad Akhtar, Tanmoy Chakraborty

Memes are powerful means for effective communication on social media. Their effortless amalgamation of viral visuals and compelling messages can have far-reaching implications with proper marketing. Previous research on memes has primarily focused on characterizing their affective spectrum and detecting whether the meme's message insinuates any intended harm, such as hate, offense, racism, etc. However, memes often use abstraction, which can be elusive. Here, we introduce a novel task - EXCLAIM, generating explanations for visual semantic role labeling in memes. To this end, we curate ExHVV, a novel dataset that offers natural language explanations of connotative roles for three types of entities - heroes, villains, and victims, encompassing 4,680 entities present in 3K memes. We also benchmark ExHVV with several strong unimodal and multimodal baselines. Moreover, we posit LUMEN, a novel multimodal, multi-task learning framework that endeavors to address EXCLAIM optimally by jointly learning to predict the correct semantic roles and correspondingly to generate suitable natural language explanations. LUMEN distinctly outperforms the best baseline across 18 standard natural language generation evaluation metrics. Our systematic evaluation and analyses demonstrate that characteristic multimodal cues required for adjudicating semantic roles are also helpful for generating suitable explanations.

📄 PDF Abstract BibTeX arXiv:2212.00715

Code (2)

LCS2-IIITD/LUMEN-Explaining-Memes 공식 구현 pytorch
dreamh1gh/awesome-srl paddle

Tasks

MarketingMulti-Task LearningSemantic Role LabelingText Generation

Similar Papers 제목 키워드 기반

Meme-ingful Analysis: Enhanced Understanding of Cyberbullying in Memes Through Multimodal Explanations

2024-01-18 · Prince Jha, Krishanu Maity, Raghav Jain, Apoorv Verma 외

Internet memes have gained significant influence in communicating political, psychological, and sociocultural ideas. While memes are often humorous, there has been a rise in the use of memes for trolling and cyberbullyin…

MemeBench: What LVLMs Miss When Interpreting Culture-Dependent Memes

2026-07-30 · Weihang Wang, Kainan Tu, Jielei Zhang, Run Yang 외 arxiv

Large vision-language models have improved at describing visual content, but accurate descriptions do not ensure interpretation when meaning depends on knowledge beyond the pixels. Memes expose this gap because they rely…

MemeMQA: Multimodal Question Answering for Memes via Rationale-Based Inferencing

2024-05-18 · Siddhant Agarwal, Shivam Sharma, Preslav Nakov, Tanmoy Chakraborty

Memes have evolved as a prevalent medium for diverse communication, ranging from humour to propaganda. With the rising popularity of image-focused content, there is a growing need to explore its potential harm from diffe…

Question AnsweringText Generation

Meme Similarity and Emotion Detection using Multimodal Analysis

2025-03-21 · Aidos Konyspay, Pakizar Shamoi, Malika Ziyada, Zhusup Smambayev

Internet memes are a central element of online culture, blending images and text. While substantial research has focused on either the visual or textual components of memes, little attention has been given to their inter…

Adapting Reinforcement Learning with Chain-of-Thought Supervision for Explainable Detection of Hateful and Propagandistic Memes

2026-06-13 · Mohamed Bayan Kmainasi, Mucahid Kutlu, Ali Ezzat Shahroor, Abul Hasnat 외 arxiv

Hateful and propagandistic memes exploit the interplay between images and text to convey harmful intent that neither modality reveals alone. Although thinking-based multimodal large language models (MLLMs) have advanced …

Reinforcement Learning