paper-with-me

Papers

MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms

2024-02-21 · Yiqiao Jin, MinJe Choi, Gaurav Verma, Jindong Wang, Srijan Kumar

Social media platforms are hubs for multimodal information exchange, encompassing text, images, and videos, making it challenging for machines to comprehend the information or emotions associated with interactions in online spaces. Multimodal Large Language Models (MLLMs) have emerged as a promising solution to these challenges, yet they struggle to accurately interpret human emotions and complex content such as misinformation. This paper introduces MM-Soc, a comprehensive benchmark designed to evaluate MLLMs' understanding of multimodal social media content. MM-Soc compiles prominent multimodal datasets and incorporates a novel large-scale YouTube tagging dataset, targeting a range of tasks from misinformation detection, hate speech detection, and social context generation. Through our exhaustive evaluation on ten size-variants of four open-source MLLMs, we have identified significant performance disparities, highlighting the need for advancements in models' social understanding capabilities. Our analysis reveals that, in a zero-shot setting, various types of MLLMs generally exhibit difficulties in handling social media tasks. However, MLLMs demonstrate performance improvements post fine-tuning, suggesting potential pathways for improvement. Our code and data are available at https://github.com/claws-lab/MMSoc.git.

📄 PDF Abstract BibTeX arXiv:2402.14154

Code (1)

claws-lab/mmsoc 공식 구현 pytorch

Tasks

BenchmarkingHate Speech DetectionMisinformation

Similar Papers 제목 키워드 기반

SNS-Bench-VL: Benchmarking Multimodal Large Language Models in Social Networking Services

2025-05-29 · Hongcheng Guo, Zheyong Xie, Shaosheng Cao, Boyang Wang 외

With the increasing integration of visual and textual content in Social Networking Services (SNS), evaluating the multimodal capabilities of Large Language Models (LLMs) is crucial for enhancing user experience, content …

BenchmarkingInformation RetrievalMultiple-choice

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

2026-06-25 · Qinkai Zhang, Yanyan Zhao, Xin Lu, Yulin Hu 외 arxiv

Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogue? More comprehensive personalization demands a harder capability --…

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection

2026-05-02 · Tianxiao Li, Zhenglin Huang, Haiquan Wen, Yiwei He 외 arxiv

Multimodal deepfakes are proliferating on social media and threaten authenticity, information integrity, and digital forensics. Existing benchmarks are constrained by their single-modality scope, simplified manipulations…

DeepFake Detection

SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge

2024-05-17 · Bo Wu, Peiye Liu, Wen-Huang Cheng, Bei Liu 외

Social Media Popularity Prediction (SMPP) is a crucial task that involves automatically predicting future popularity values of online posts, leveraging vast amounts of multimodal data available on social media platforms.…

BenchmarkingSocial Media Popularity Prediction

SoMeLVLM: A Large Vision Language Model for Social Media Processing

2024-02-20 · Xinnong Zhang, Haoyu Kuang, Xinyi Mou, Hanjia Lyu 외

The growth of social media, characterized by its multimodal nature, has led to the emergence of diverse phenomena and challenges, which calls for an effective approach to uniformly solve automated tasks. The powerful Lar…

Language ModelingLanguage Modelling