VALD: Multi-Stage Vision Attack Detection for Efficient LVLM Defense
Large Vision-Language Models (LVLMs) can be vulnerable to adversarial images that subtly bias their outputs toward plausible yet incorrect responses. We introduce a general, efficient, and training-free defense that combines image transformations with agentic data consolidation to recover correct model behavior. A key component of our approach is a two-stage detection mechanism that quickly filters out the majority of clean inputs. We first assess image consistency under content-preserving transformations at negligible computational cost. For more challenging cases, we examine discrepancies in a text-embedding space. Only when necessary do we invoke a powerful LLM to resolve attack-induced divergences. A key idea is to consolidate multiple responses, leveraging both their similarities and their differences. We show that our method achieves state-of-the-art accuracy while maintaining notable efficiency: most clean images skip costly processing, and even in the presence of numerous adversarial examples, the overhead remains minimal.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
EvaLDA: Efficient Evasion Attacks Towards Latent Dirichlet Allocation
As one of the most powerful topic models, Latent Dirichlet Allocation (LDA) has been used in a vast range of tasks, including document understanding, information retrieval and peer-reviewer assignment. Despite its tremen…
document understandingInformation RetrievalRetrievalSentiment Analysis+1MixMicrobleed: Multi-stage detection and segmentation of cerebral microbleeds
Cerebral microbleeds are small, dark, round lesions that can be visualised on T2*-weighted MRI or other sequences sensitive to susceptibility effects. In this work, we propose a multi-stage approach to both microbleed de…
SegmentationPseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions
Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy termed "value-action gap." In this work, we argue that this gap persist…
VALD-GAN: video anomaly detection using latent discriminator augmented GAN
The most crucial and difficult challenge for intelligent video surveillance is to identify anomalies in a video that comprises anomalous behavior or occurrences. The ambiguous definition of the anomaly makes the detectio…
Anomaly DetectionAnomaly Detection In Surveillance VideosVideo Anomaly DetectionIntroducing EVALD -- Software Applications for Automatic Evaluation of Discourse in Czech
In the paper, we introduce two software applications for automatic evaluation of coherence in Czech texts called EVALD {--} Evaluator of Discourse. The first one {--} EVALD 1.0 {--} evaluates texts written by native spea…