paper-with-me

홈 › Papers

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

2026-03-23 · Xinghan Li, Junhao Xu, Jingjing Chen arxiv

Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However, the reasoning process of current MLLM-based methods combines evidence generation and manipulation localization into a unified step. This combination blurs the boundary between faithful observations and hallucinated explanations, leading to unreliable conclusions. Building on this, we present VIGIL, a part-centric structured forensic framework inspired by expert forensic practice through a plan-then-examine pipeline: the model first plans which facial parts warrant inspection based on global visual cues, then examines each part with independently sourced forensic evidence. A stage-gated injection mechanism delivers part-level forensic evidence only during examination, ensuring that part selection remains driven by the model's own perception rather than biased by external signals. We further propose a progressive three-stage training paradigm whose reinforcement learning stage employs part-aware rewards to enforce anatomical validity and evidence--conclusion coherence. To enable rigorous generalizability evaluation, we construct OmniFake, a hierarchical 5-Level benchmark where the model, trained on only three foundational generators, is progressively tested up to in-the-wild social-media data. Extensive experiments on OmniFake and cross-dataset evaluations demonstrate that VIGIL consistently outperforms both expert detectors and concurrent MLLM-based methods across all generalizability levels.

📄 PDF Abstract BibTeX arXiv:2603.21526

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningDeepFake Detection

Similar Papers 제목 키워드 기반

VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit

2026-01-09 · Junda Lin, Zhaomeng Zhou, Zhi Zheng, Shuochen Liu 외 arxiv

LLM agents operating in open environments face escalating risks from indirect prompt injection, particularly within the tool stream where manipulated metadata and runtime feedback hijack execution flow. Existing defenses…

Beyond Skepticism: Evaluating LLMs Pedagogical Intent Reasoning with the Adaptive Pedagogical Vigilance Framework

2026-07-02 · Minghao Chen, Ruihan Zhou, Jiayi Tang, Zihan Xu 외 arxiv

The capacity of Large Language Models (LLMs) to reason about pedagogical intent within instructional communication remains underexplored, particularly in educational domains such as translation pedagogy. To address this,…

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

2026-05-08 · Manish Bhattarai, Ismael Boureima, Nishath Rajiv Ranasinghe, Scott Pakin 외 arxiv

We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal: instead of a binary outcome or a single holistic score, each respons…

Reinforcement Learning

Critical appraisal of artificial intelligence for rare-event recognition: principles and pharmacovigilance case studies

2025-10-05 · G. Niklas Noren, Eva-Lisa Meldau, Johan Ellenius arxiv

Many high-stakes AI applications target low-prevalence events, where apparent accuracy can conceal limited real-world value. Relevant AI models range from expert-defined rules and traditional machine learning to generati…

Scan, Materialize, Simulate: A Generalizable Framework for Physically Grounded Robot Planning

2025-05-20 · Amine Elhafsi, Daniel Morton, Marco Pavone

Autonomous robots must reason about the physical consequences of their actions to operate effectively in unstructured, real-world environments. We present Scan, Materialize, Simulate (SMS), a unified framework that combi…

Semantic Segmentation