paper-with-me

홈 › Papers

TriDF: Evaluating Perception, Detection, and Hallucination for Interpretable DeepFake Detection

2025-12-11 · Jian-Yu Jiang-Lin, Kang-Yang Huang, Ling Zou, Ling Lo, Sheng-Ping Yang, Yu-Wen Tseng, Kun-Hsiang Lin, Chia-Ling Chen, Yu-Ting Ta, Yan-Tsung Wang, Po-Ching Chen, Hongxia Xie, Hong-Han Shuai, Wen-Huang Cheng arxiv

Advances in generative modeling have made it increasingly easy to fabricate realistic portrayals of individuals, creating serious risks for security, communication, and public trust. Detecting such person-driven manipulations requires systems that not only distinguish altered content from authentic media but also provide clear and reliable reasoning. In this paper, we introduce TriDF, a comprehensive benchmark for interpretable DeepFake detection. TriDF contains high-quality forgeries from advanced synthesis models, covering 16 DeepFake types across image, video, and audio modalities. The benchmark evaluates three key aspects: Perception, which measures the ability of a model to identify fine-grained manipulation artifacts using human-annotated evidence; Detection, which assesses classification performance across diverse forgery families and generators; and Hallucination, which quantifies the reliability of model-generated explanations. Experiments on state-of-the-art multimodal large language models show that accurate perception is essential for reliable detection, but hallucination can severely disrupt decision-making, revealing the interdependence of these three aspects. TriDF provides a unified framework for understanding the interaction between detection accuracy, evidence identification, and explanation reliability, offering a foundation for building trustworthy systems that address real-world synthetic media threats.

📄 PDF Abstract BibTeX arXiv:2512.10652

Code (0)

등록된 구현이 없습니다.

Tasks

DeepFake Detection

Similar Papers 제목 키워드 기반

EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models

2025-05-16 · Bohao Xing, Xin Liu, Guoying Zhao, Chengyu Liu 외

Emotion understanding is a critical yet challenging task. Recent advances in Multimodal Large Language Models (MLLMs) have significantly enhanced their capabilities in this area. However, MLLMs often suffer from hallucin…

Hallucination

ReactBench: A Cause-Driven Benchmark for Multimodal Hallucination via Systematic Evaluation

2026-05-28 · Shizhe Zhou, Bohan Jia, Kai Wu, Yan Shen 외 arxiv

While multimodal large language models (MLLMs) have achieved rapid progress in vision-language understanding, they remain prone to multimodal hallucinations, producing responses that are inconsistent with the visual inpu…

Evaluating and Enhancing Trustworthiness of LLMs in Perception Tasks

2024-07-18 · Malsha Ashani Mahawatta Dona, Beatriz Cabrero-Daniel, Yinan Yu, Christian Berger

Today's advanced driver assistance systems (ADAS), like adaptive cruise control or rear collision warning, are finding broader adoption across vehicle classes. Integrating such advanced, multimodal Large Language Models …

Hallucinationobject-detectionObject DetectionObject Localization+1

EPSM: A Novel Metric to Evaluate the Safety of Environmental Perception in Autonomous Driving

2025-12-17 · Jörg Gamerdinger, Sven Teufel, Stephan Amann, Lukas Marc Listl 외 arxiv

Extensive evaluation of perception systems is crucial for ensuring the safety of intelligent vehicles in complex driving scenarios. Conventional performance metrics such as precision, recall and the F1-score assess the o…

Autonomous DrivingObject DetectionLane Detection

Fakes of Varying Shades: How Warning Affects Human Perception and Engagement Regarding LLM Hallucinations

2024-04-04 · Mahjabin Nahar, Haeseung Seo, Eun-Ju Lee, Aiping Xiong 외

The widespread adoption and transformative effects of large language models (LLMs) have sparked concerns regarding their capacity to produce inaccurate and fictitious content, referred to as `hallucinations'. Given the p…

HallucinationHuman DetectionSurvey