paper-with-me

Papers

VMID: A Multimodal Fusion LLM Framework for Detecting and Identifying Misinformation of Short Videos

2024-11-15 · Weihao Zhong, Yinhao Xiao, Minghui Xu, Xiuzhen Cheng

Short video platforms have become important channels for news dissemination, offering a highly engaging and immediate way for users to access current events and share information. However, these platforms have also emerged as significant conduits for the rapid spread of misinformation, as fake news and rumors can leverage the visual appeal and wide reach of short videos to circulate extensively among audiences. Existing fake news detection methods mainly rely on single-modal information, such as text or images, or apply only basic fusion techniques, limiting their ability to handle the complex, multi-layered information inherent in short videos. To address these limitations, this paper presents a novel fake news detection method based on multimodal information, designed to identify misinformation through a multi-level analysis of video content. This approach effectively utilizes different modal representations to generate a unified textual description, which is then fed into a large language model for comprehensive evaluation. The proposed framework successfully integrates multimodal features within videos, significantly enhancing the accuracy and reliability of fake news detection. Experimental results demonstrate that the proposed approach outperforms existing models in terms of accuracy, robustness, and utilization of multimodal information, achieving an accuracy of 90.93%, which is significantly higher than the best baseline model (SV-FEND) at 81.05%. Furthermore, case studies provide additional evidence of the effectiveness of the approach in accurately distinguishing between fake news, debunking content, and real incidents, highlighting its reliability and robustness in real-world applications.

📄 PDF Abstract BibTeX arXiv:2411.10032

Code (0)

등록된 구현이 없습니다.

Tasks

Fake News DetectionLarge Language ModelMisinformation

Similar Papers 제목 키워드 기반

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion

2025-05-17 · Yinghui Zhang, Tailin Chen, Yuchen Zhang, Zeyu Fu

The rapid rise of video content on platforms such as TikTok and YouTube has transformed information dissemination, but it has also facilitated the spread of harmful content, particularly hate videos. Despite significant …

Multimodal Learning for Fake News Detection in Short Videos Using Linguistically Verified Data and Heterogeneous Modality Fusion

2025-09-19 · Shanghong Li, Chiam Wen Qi Ruth, Hong Xu, Fang Liu arxiv

The rapid proliferation of short video platforms has necessitated advanced methods for detecting fake news. This need arises from the widespread influence and ease of sharing misinformation, which can lead to significant…

Fake News Detection

BigTokDetect: A Clinically-Informed Vision-Language Modeling Framework for Detecting Pro-Bigorexia Videos on TikTok

2025-07-30 · Minh Duc Chu, Kshitij Pawar, Zihao He, Roxanna Sharifi 외 arxiv

Social media platforms face escalating challenges in detecting harmful content that promotes muscle dysmorphic behaviors and cognitions (bigorexia). This content can evade moderation by camouflaging as legitimate fitness…

MOMENTA: A Multimodal Framework for Detecting Harmful Memes and Their Targets

2021-09-11 · Findings (EMNLP) 2021 11 · Shraman Pramanick, Shivam Sharma, Dimitar Dimitrov, Md Shad Akhtar 외

Internet memes have become powerful means to transmit political, psychological, and socio-cultural ideas. Although memes are typically humorous, recent days have witnessed an escalation of harmful memes used for trolling…

MixMAS: A Framework for Sampling-Based Mixer Architecture Search for Multimodal Fusion and Learning

2024-12-24 · Abdelmadjid Chergui, Grigor Bezirganyan, Sana Sellami, Laure Berti-ÉQuille 외

Choosing a suitable deep learning architecture for multimodal data fusion is a challenging task, as it requires the effective integration and processing of diverse data types, each with distinct structures and characteri…

Benchmarking