paper-with-me

홈 › Papers

Misinformation Span Detection in Videos via Audio Transcripts

2026-04-23 · Breno Matos, Rennan C. Lima, Savvas Zannettou, Fabricio Benevenuto, Rodrygo L. T. Santos arxiv

Online misinformation is one of the most challenging issues lately, yielding severe consequences, including political polarization, attacks on democracy, and public health risks. Misinformation manifests in any platform with a large user base, including online social networks and messaging apps. It permeates all media and content forms, including images, text, audio, and video. Distinctly, video-based misinformation represents a multifaceted challenge for fact-checkers, given the ease with which individuals can record and upload videos on various video-sharing platforms. Previous research efforts investigated detecting video-based misinformation, focusing on whether a video shares misinformation or not on a video level. While this approach is useful, it only provides a limited and non-easily interpretable view of the problem given that it does not provide an additional context of when misinformation occurs within videos and what content (i.e., claims) are responsible for the video's misinformation nature. In this work, we attempt to bridge this research gap by creating two novel datasets that allow us to explore misinformation detection on videos via audio transcripts, focusing on identifying the span of videos that are responsible for the video's misinformation claim (misinformation span detection). We present two new datasets for this task. We transcribe each video's audio to text, identifying the video segment in which the misinformation claims appears, resulting in two datasets of more than 500 videos with over 2,400 segments containing annotated fact-checked claims. Then, we employ classifiers built with state-of-the-art language models, and our results show that we can identify in which part of a video there is misinformation with an F1 score of 0.68. We make publicly available our annotated datasets. We also release all transcripts, audio and videos.

📄 PDF Abstract BibTeX arXiv:2604.21767

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ViClaim: A Multilingual Multilabel Dataset for Automatic Claim Detection in Videos

2025-04-17 · Patrick Giedemann, Pius von Däniken, Jan Deriu, Alvaro Rodrigo 외

The growing influence of video content as a medium for communication and misinformation underscores the urgent need for effective tools to analyze claims in multilingual and multi-topic settings. Existing efforts in misi…

MisinformationSentence

Human Detection of Political Speech Deepfakes across Transcripts, Audio, and Video

2022-02-25 · Matthew Groh, Aruna Sankaranarayanan, Nikhil Singh, Dong Young Kim 외

Recent advances in technology for hyper-realistic visual and audio effects provoke the concern that deepfake videos of political speeches will soon be indistinguishable from authentic video recordings. The conventional w…

Face SwappingHuman DetectionMisinformationtext-to-speech+1

Poster: Exploring the Limits of Audio-Based Detection of Turkish Phone Call Scams

2026-06-23 · Arda Eren, Micheal Cheung, Youqian Zhang, Grace Ngai 외 arxiv

Scam phone calls exploit vulnerable communities worldwide, yet research on detection has focused almost exclusively on English and other high-resource languages. In low-resource settings such as Turkish, detection is esp…

When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

2026-04-18 · Chaewan Chun, Delvin Ce Zhang, Dongwon Lee arxiv

Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and live streams. With millions of shows and hundreds of millions of listen…

METER: Multi-modal Evidence-based Thinking and Explainable Reasoning -- Algorithm and Benchmark

2025-07-22 · Xu Yang, Qi Zhang, Shuming Jiang, Yaowen Xu 외 arxiv

With the rapid advancement of generative AI, synthetic content across images, videos, and audio has become increasingly realistic, amplifying the risk of misinformation. Existing detection approaches predominantly focus …

Binary Classification