paper-with-me

Papers

AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM

2025-03-06 · Sunghyun Ahn, Youngwan Jo, Kijung Lee, Sein Kwon, Inpyo Hong, Sanghyun Park

Video anomaly detection (VAD) is crucial for video analysis and surveillance in computer vision. However, existing VAD models rely on learned normal patterns, which makes them difficult to apply to diverse environments. Consequently, users should retrain models or develop separate AI models for new environments, which requires expertise in machine learning, high-performance hardware, and extensive data collection, limiting the practical usability of VAD. To address these challenges, this study proposes customizable video anomaly detection (C-VAD) technique and the AnyAnomaly model. C-VAD considers user-defined text as an abnormal event and detects frames containing a specified event in a video. We effectively implemented AnyAnomaly using a context-aware visual question answering without fine-tuning the large vision language model. To validate the effectiveness of the proposed model, we constructed C-VAD datasets and demonstrated the superiority of AnyAnomaly. Furthermore, our approach showed competitive performance on VAD benchmark datasets, achieving state-of-the-art results on the UBnormal dataset and outperforming other methods in generalization across all datasets. Our code is available online at github.com/SkiddieAhn/Paper-AnyAnomaly.

📄 PDF Abstract BibTeX arXiv:2503.04504

Code (1)

SkiddieAhn/Paper-AnyAnomaly 공식 구현 pytorch

Tasks

Anomaly DetectionLanguage ModelingLanguage ModellingQuestion AnsweringVideo Anomaly DetectionVisual Question Answering

Similar Papers 제목 키워드 기반

A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis

2025-11-02 · Dongheng Lin, Mengxue Qu, Kunyang Han, Jianbo Jiao 외 arxiv

Most video-anomaly research stops at frame-wise detection, offering little insight into why an event is abnormal, typically outputting only frame-wise anomaly scores without spatial or semantic context. Recent video anom…

Video Anomaly Detection

No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection

2026-02-22 · Zunkai Dai, Ke Li, Jiajia Liu, Jie Yang 외 arxiv

The collection and detection of video anomaly data has long been a challenging problem due to its rare occurrence and spatio-temporal scarcity. Existing video anomaly detection (VAD) methods under perform in open-world s…

Video Anomaly Detection

Flashback: Memory-Driven Zero-shot, Real-time Video Anomaly Detection

2025-05-21 · Hyogun Lee, Haksub Kim, Ig-Jae Kim, Yonghun Choi

Video Anomaly Detection (VAD) automatically identifies anomalous events from video, mitigating the need for human operators in large-scale surveillance deployments. However, three fundamental obstacles hinder real-world …

Anomaly DetectionGPUVideo Anomaly Detection

TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection

2025-11-01 · Yousuf Ahmed Siddiqui, Sufiyaan Usmani, Umer Tariq, Jawwad Ahmed Shamsi 외 arxiv

Video anomalies often depend on contextual information available and temporal evolution. Non-anomalous action in one context can be anomalous in some other context. Most anomaly detectors, however, do not notice this typ…

Video Anomaly DetectionAnomaly Classification

TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis

2025-05-20 · Yu Zhang, Wenxiang Guo, Changhao Pan, Dongyu Yao 외

Customizable multilingual zero-shot singing voice synthesis (SVS) has various potential applications in music composition and short video dubbing. However, existing SVS models overly depend on phoneme and note boundary a…

Contrastive LearningSinging Voice SynthesisStyle Transfer