paper-with-me

홈 › Papers

Disrupting Hierarchical Reasoning: Adversarial Protection for Geographic Privacy in Multimodal Reasoning Models

2025-12-09 · Jiaming Zhang, Che Wang, Yang Cao, Longtao Huang, Wei Yang Bryan Lim arxiv

Multi-modal large reasoning models (MLRMs) pose significant privacy risks by inferring precise geographic locations from personal images through hierarchical chain-of-thought reasoning. Existing privacy protection techniques, primarily designed for perception-based models, prove ineffective against MLRMs' sophisticated multi-step reasoning processes that analyze environmental cues. We introduce \textbf{ReasonBreak}, a novel adversarial framework specifically designed to disrupt hierarchical reasoning in MLRMs through concept-aware perturbations. Our approach is founded on the key insight that effective disruption of geographic reasoning requires perturbations aligned with conceptual hierarchies rather than uniform noise. ReasonBreak strategically targets critical conceptual dependencies within reasoning chains, generating perturbations that invalidate specific inference steps and cascade through subsequent reasoning stages. To facilitate this approach, we contribute \textbf{GeoPrivacy-6K}, a comprehensive dataset comprising 6,341 ultra-high-resolution images ($\geq$2K) with hierarchical concept annotations. Extensive evaluation across seven state-of-the-art MLRMs (including GPT-o3, GPT-5, Gemini 2.5 Pro) demonstrates ReasonBreak's superior effectiveness, achieving a 14.4\% improvement in tract-level protection (33.8\% vs 19.4\%) and nearly doubling block-level protection (33.5\% vs 16.8\%). This work establishes a new paradigm for privacy protection against reasoning-based threats.

📄 PDF Abstract BibTeX arXiv:2512.08503

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Reasoning

Similar Papers 제목 키워드 기반

Local Features Meet Stochastic Anonymization: Revolutionizing Privacy-Preserving Face Recognition for Black-Box Models

2024-12-11 · Yuanwei Liu, Chengyu Jia, Ruqi Xiao, Xuemai Jia 외

The task of privacy-preserving face recognition (PPFR) currently faces two major unsolved challenges: (1) existing methods are typically effective only on specific face recognition models and struggle to generalize to bl…

Face RecognitionPrivacy Preserving

GeoShield: Safeguarding Geolocation Privacy from Vision-Language Models via Adversarial Perturbations

2025-08-05 · Xinwei Liu, Xiaojun Jia, Yuan Xun, Simeng Qin 외 arxiv

Vision-Language Models (VLMs) such as GPT-4o now demonstrate a remarkable ability to infer users' locations from public shared images, posing a substantial risk to geoprivacy. Although adversarial perturbations offer a p…

Beyond Pixels: Semantic-aware Typographic Attack for Geo-Privacy Protection

2025-11-16 · Jiayi Zhu, Yihao Huang, Yue Cao, Xiaojun Jia 외 arxiv

Large Visual Language Models (LVLMs) now pose a serious yet overlooked privacy threat, as they can infer a social media user's geolocation directly from shared images, leading to unintended privacy leakage. While adversa…

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models

2026-07-28 · Yimao Guo, Zuomin Qu, Wei Lu arxiv

The rapid advancement of video generation models has led to the increasing misuse of image-to-video (I2V) models. Although substantial progress has been made in detecting AI-generated videos, proactive defenses against I…

Video Generation

DePrompt: Desensitization and Evaluation of Personal Identifiable Information in Large Language Model Prompts

2024-08-16 · Xiongtao Sun, Gan Liu, Zhipeng He, Hui Li 외

Prompt serves as a crucial link in interacting with large language models (LLMs), widely impacting the accuracy and interpretability of model outputs. However, acquiring accurate and high-quality responses necessitates p…

Language ModelingLanguage ModellingLarge Language Model