paper-with-me

Papers

Cognitive Inception: Agentic Reasoning against Visual Deceptions by Injecting Skepticism

2025-11-21 · Yinjie Zhao, Heng Zhao, Bihan Wen, Joey Tianyi Zhou arxiv

As the development of AI-generated contents (AIGC), multi-modal Large Language Models (LLM) struggle to identify generated visual inputs from real ones. Such shortcoming causes vulnerability against visual deceptions, where the models are deceived by generated contents, and the reliability of reasoning processes is jeopardized. Therefore, facing rapidly emerging generative models and diverse data distribution, it is of vital importance to improve LLMs' generalizable reasoning to verify the authenticity of visual inputs against potential deceptions. Inspired by human cognitive processes, we discovered that LLMs exhibit tendency of over-trusting the visual inputs, while injecting skepticism could significantly improve the models visual cognitive capability against visual deceptions. Based on this discovery, we propose \textbf{Inception}, a fully reasoning-based agentic reasoning framework to conduct generalizable authenticity verification by injecting skepticism, where LLMs' reasoning logic is iteratively enhanced between External Skeptic and Internal Skeptic agents. To the best of our knowledge, this is the first fully reasoning-based framework against AIGC visual deceptions. Our approach achieved a large margin of performance improvement over the strongest existing LLM baselines and SOTA performance on AEGIS benchmark.

📄 PDF Abstract BibTeX arXiv:2511.17672

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Visual Inception: Compromising Long-term Planning in Agentic Recommenders via Multimodal Memory Poisoning

2026-04-18 · Jiachen Qian arxiv

The evolution from static ranking models to Agentic Recommender Systems (Agentic RecSys) empowers AI agents to maintain long-term user profiles and autonomously plan service tasks. While this paradigm shift enhances pers…

Evaluating Cognitive Age Alignment in Interactive AI Agents

2026-05-18 · Yifan Shen, Jiawen Zhang, Jian Xu, Junho Kim 외 arxiv

While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across domains ranging from daily life to advanced scientific research, a profo…

Visual Reasoning

Agentic AI Enhances Physician Trust in Clinical Decision Making

2026-06-16 · Zhiling Yan, Zhe Fang, David J King, Ann Pongsakul 외 arxiv

Medical AI has shifted from reasoning to agentic AI, a new paradigm that autonomously invokes external tools during reasoning, rendering intermediate reasoning steps and tool outputs transparent to users. Although proven…

Decision Making

I2I-STRADA -- Information to Insights via Structured Reasoning Agent for Data Analysis

2025-07-23 · SaiBarath Sundar, Pranav Satheesan, Udayaadithya Avadhanam arxiv

Recent advances in agentic systems for data analysis have emphasized automation of insight generation through multi-agent frameworks, and orchestration layers. While these systems effectively manage tasks like query tran…

OpenQlaw: An Agentic AI Assistant for Analysis of 2D Quantum Materials

2026-03-17 · Sankalp Pandey, Xuan-Bac Nguyen, Hoang-Quan Nguyen, Tim Faltermeier 외 arxiv

The transition from optical identification of 2D quantum materials to practical device fabrication requires dynamic reasoning beyond the detection accuracy. While recent domain-specific Multimodal Large Language Models (…