paper-with-me

홈 › Papers

VLDBench: Vision Language Models Disinformation Detection Benchmark

2025-02-17 · Shaina Raza, Ashmal Vayani, Aditya Jain, Aravind Narayanan, Vahid Reza Khazaie, Syed Raza Bashir, Elham Dolatabadi, Gias Uddin, Christos Emmanouilidis, Rizwan Qureshi, Mubarak Shah

The rapid rise of AI-generated content has made detecting disinformation increasingly challenging. In particular, multimodal disinformation, i.e., online posts-articles that contain images and texts with fabricated information are specially designed to deceive. While existing AI safety benchmarks primarily address bias and toxicity, multimodal disinformation detection remains largely underexplored. To address this challenge, we present the Vision-Language Disinformation Detection Benchmark VLDBench, the first comprehensive benchmark for detecting disinformation across both unimodal (text-only) and multimodal (text and image) content, comprising 31,000} news article-image pairs, spanning 13 distinct categories, for robust evaluation. VLDBench features a rigorous semi-automated data curation pipeline, with 22 domain experts dedicating 300 plus hours} to annotation, achieving a strong inter-annotator agreement (Cohen kappa = 0.78). We extensively evaluate state-of-the-art Large Language Models (LLMs) and Vision-Language Models (VLMs), demonstrating that integrating textual and visual cues in multimodal news posts improves disinformation detection accuracy by 5 - 35 % compared to unimodal models. Developed in alignment with AI governance frameworks such as the EU AI Act, NIST guidelines, and the MIT AI Risk Repository 2024, VLDBench is expected to become a benchmark for detecting disinformation in online multi-modal contents. Our code and data will be publicly available.

📄 PDF Abstract BibTeX arXiv:2502.11361

Code (1)

VectorInstitute/VLDBench pytorch

Tasks

Articles

Similar Papers 제목 키워드 기반

MALicious INTent Dataset and Inoculating LLMs for Enhanced Disinformation Detection

2026-03-15 · Arkadiusz Modzelewski, Witold Sosnowski, Eleni Papadopulos, Elisa Sartori 외 arxiv

The intentional creation and spread of disinformation poses a significant threat to public discourse. However, existing English datasets and research rarely address the intentionality behind the disinformation. This work…

Intent Classification

Taking a Stance on Fake News: Towards Automatic Disinformation Assessment via Deep Bidirectional Transformer Language Models for Stance Detection

2019-11-27 · Chris Dulhanty, Jason L. Deglint, Ibrahim Ben Daya, Alexander Wong

The exponential rise of social media and digital news in the past decade has had the unfortunate consequence of escalating what the United Nations has called a global topic of concern: the growing prevalence of disinform…

ArticlesLanguage ModelingLanguage ModellingStance Detection+2

FANG-COVID: A New Large-Scale Benchmark Dataset for Fake News Detection in German

2021-11-01 · EMNLP (FEVER) 2021 11 · Justus Mattern, Yu Qiao, Elma Kerz, Daniel Wiechmann 외

As the world continues to fight the COVID-19 pandemic, it is simultaneously fighting an ‘infodemic’ – a flood of disinformation and spread of conspiracy theories leading to health threats and the division of society. To …

ArticlesFake News Detection

Multimodal Climate Disinformation Detection: Integrating Vision-Language Models with External Knowledge Sources

2026-01-22 · Marzieh Adeli Shamsabad, Hamed Ghodrati arxiv

Climate disinformation has become a major challenge in today digital world, especially with the rise of misleading images and videos shared widely on social media. These false claims are often convincing and difficult to…

Disinformation Detection: An Evolving Challenge in the Age of LLMs

2023-09-25 · Bohan Jiang, Zhen Tan, Ayushi Nirmal, Huan Liu

The advent of generative Large Language Models (LLMs) such as ChatGPT has catalyzed transformative advancements across multiple domains. However, alongside these advancements, they have also introduced potential threats.…