paper-with-me

홈 › Papers

VALD: Multi-Stage Vision Attack Detection for Efficient LVLM Defense

2026-02-23 · Nadav Kadvil, Malak Fares, Ayellet Tal arxiv

Large Vision-Language Models (LVLMs) can be vulnerable to adversarial images that subtly bias their outputs toward plausible yet incorrect responses. We introduce a general, efficient, and training-free defense that combines image transformations with agentic data consolidation to recover correct model behavior. A key component of our approach is a two-stage detection mechanism that quickly filters out the majority of clean inputs. We first assess image consistency under content-preserving transformations at negligible computational cost. For more challenging cases, we examine discrepancies in a text-embedding space. Only when necessary do we invoke a powerful LLM to resolve attack-induced divergences. A key idea is to consolidate multiple responses, leveraging both their similarities and their differences. We show that our method achieves state-of-the-art accuracy while maintaining notable efficiency: most clean images skip costly processing, and even in the presence of numerous adversarial examples, the overhead remains minimal.

📄 PDF Abstract BibTeX arXiv:2602.19570

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EvaLDA: Efficient Evasion Attacks Towards Latent Dirichlet Allocation

2020-12-09 · Qi Zhou, Haipeng Chen, Yitao Zheng, Zhen Wang

As one of the most powerful topic models, Latent Dirichlet Allocation (LDA) has been used in a vast range of tasks, including document understanding, information retrieval and peer-reviewer assignment. Despite its tremen…

document understandingInformation RetrievalRetrievalSentiment Analysis+1

MixMicrobleed: Multi-stage detection and segmentation of cerebral microbleeds

2021-08-05 · Marta Girones Sanguesa, Denis Kutnar, Bas H. M. van der Velden, Hugo J. Kuijf

Cerebral microbleeds are small, dark, round lesions that can be visualised on T2*-weighted MRI or other sequences sensitive to susceptibility effects. In this work, we propose a multi-stage approach to both microbleed de…

Segmentation

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

2026-05-11 · Sushrita Rakshit, Hanwen Zhang, Hua Shen arxiv

Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy termed "value-action gap." In this work, we argue that this gap persist…

VALD-GAN: video anomaly detection using latent discriminator augmented GAN

2023-10-18 · Signal, Image and Video Processing 2023 10 · Rituraj Singh, Anikeit Sethi, Krishanu Saini, Sumeet Saurav 외

The most crucial and difficult challenge for intelligent video surveillance is to identify anomalies in a video that comprises anomalous behavior or occurrences. The ambiguous definition of the anomaly makes the detectio…

Anomaly DetectionAnomaly Detection In Surveillance VideosVideo Anomaly Detection

Introducing EVALD -- Software Applications for Automatic Evaluation of Discourse in Czech

2017-09-01 · RANLP 2017 9 · Kate{\v{r}}ina Rysov{\'a}, Magdal{\'e}na Rysov{\'a}, Ji{\v{r}}{\'\i} M{\'\i}rovsk{\'y}, Michal Nov{\'a}k

In the paper, we introduce two software applications for automatic evaluation of coherence in Czech texts called EVALD {--} Evaluator of Discourse. The first one {--} EVALD 1.0 {--} evaluates texts written by native spea…