paper-with-me

홈 › Papers

M4FC: a Multimodal, Multilingual, Multicultural, Multitask Real-World Fact-Checking Dataset

2025-10-27 · Jiahui Geng, Jonathan Tonglet, Iryna Gurevych arxiv

Existing real-world datasets for multimodal fact-checking have multiple limitations: they contain few instances, cover on only one or two languages, focus only on one task, or rely on external news article sets for sourcing true claims. To address these shortcomings, we introduce M4FC, a new real-world dataset comprising 4,982 images paired with 6,980 claims. The images, verified by professional fact-checkers from 22 organizations, represent a diverse range of cultural and geographic contexts. Each claim is available in one or two out of ten languages. M4FC spans six multimodal fact-checking tasks: visual claim extraction, claimant intent prediction, fake image detection, image contextualization, location verification, and verdict prediction. We provide baseline results for all tasks and analyze how combining intermediate tasks affects verdict prediction performance. We make our dataset and code publicly available.

📄 PDF Abstract BibTeX arXiv:2510.23508

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Culture In a Frame: C$^3$B as a Comic-Based Benchmark for Multimodal Culturally Awareness

2025-09-27 · Yuchen Song, Andong Chen, Wenxin Zhu, Kehai Chen 외 arxiv

Cultural awareness capabilities have emerged as a critical capability for Multimodal Large Language Models (MLLMs). However, current benchmarks lack progressed difficulty in their task design and are deficient in cross-l…

Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models

2025-01-02 · Bin Wang, Xunlong Zou, Shuo Sun, Wenyu Zhang 외

Singlish, a Creole language rooted in English, is a key focus in linguistic research within multilingual and multicultural contexts. However, its spoken form remains underexplored, limiting insights into its linguistic s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Question Answeringspeech-recognition+1

Enhancing Multilingual Information Retrieval in Mixed Human Resources Environments: A RAG Model Implementation for Multicultural Enterprise

2024-01-03 · Syed Rameel Ahmad

The advent of Large Language Models has revolutionized information retrieval, ushering in a new era of expansive knowledge accessibility. While these models excel in providing open-world knowledge, effectively extracting…

Information RetrievalRAGRetrievalRetrieval-augmented Generation+1

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

2026-04-21 · Peiqin Lin, Chenyang Lyu, Wenjiang Luo, Haotian Ye 외 arxiv

Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities. However, these benchmarks prioritize generic language understanding or…

Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation

2025-04-09 · Israfel Salazar, Manuel Fernández Burda, Shayekh Bin Islam, Arshia Soltani Moakhar 외

The evaluation of vision-language models (VLMs) has mainly relied on English-language benchmarks, leaving significant gaps in both multilingual and multicultural coverage. While multilingual benchmarks have expanded, bot…

Multiple-choice