paper-with-me

홈 › Papers

Just Dance with $π$! A Poly-modal Inductor for Weakly-supervised Video Anomaly Detection

2025-05-19 · Snehashis Majhi, Giacomo D'Amicantonio, Antitza Dantcheva, Quan Kong, Lorenzo Garattoni, Gianpiero Francesca, Egor Bondarev, Francois Bremond

Weakly-supervised methods for video anomaly detection (VAD) are conventionally based merely on RGB spatio-temporal features, which continues to limit their reliability in real-world scenarios. This is due to the fact that RGB-features are not sufficiently distinctive in setting apart categories such as shoplifting from visually similar events. Therefore, towards robust complex real-world VAD, it is essential to augment RGB spatio-temporal features by additional modalities. Motivated by this, we introduce the Poly-modal Induced framework for VAD: "PI-VAD", a novel approach that augments RGB representations by five additional modalities. Specifically, the modalities include sensitivity to fine-grained motion (Pose), three dimensional scene and entity representation (Depth), surrounding objects (Panoptic masks), global motion (optical flow), as well as language cues (VLM). Each modality represents an axis of a polygon, streamlined to add salient cues to RGB. PI-VAD includes two plug-in modules, namely Pseudo-modality Generation module and Cross Modal Induction module, which generate modality-specific prototypical representation and, thereby, induce multi-modal information into RGB cues. These modules operate by performing anomaly-aware auxiliary tasks and necessitate five modality backbones -- only during training. Notably, PI-VAD achieves state-of-the-art accuracy on three prominent VAD datasets encompassing real-world scenarios, without requiring the computational overhead of five modality backbones at inference.

📄 PDF Abstract BibTeX arXiv:2505.13123

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionOptical Flow EstimationVideo Anomaly DetectionWeakly-supervised Video Anomaly Detection

Similar Papers 제목 키워드 기반

Just Dance with pi! A Poly-modal Inductor for Weakly-supervised Video Anomaly Detection

2025-01-01 · CVPR 2025 1 · Snehashis Majhi, Giacomo D'Amicantonio, Antitza Dantcheva, Quan Kong 외

Weakly-supervised methods for video anomaly detection (VAD) are conventionally based merely on RGB spatio-temporal features, which continues to limit their reliability in real-world scenarios. This is due to the fact…

Anomaly DetectionOptical Flow EstimationVideo Anomaly DetectionWeakly-supervised Video Anomaly Detection

PolyBlocks: A Compiler Infrastructure for AI Chips and Programming Frameworks

2026-03-06 · Uday Bondhugula, Akshay Baviskar, Navdeep Katel, Vimal Patel 외 arxiv

We present the design and implementation of PolyBlocks, a modular and reusable MLIR-based compiler infrastructure for AI programming frameworks and AI chips. PolyBlocks is based on pass pipelines that compose transformat…

Code Generation

Logical Induction

2016-09-12 · Scott Garrabrant, Tsvi Benson-Tilsen, Andrew Critch, Nate Soares 외

We present a computable algorithm that assigns probabilities to every logical statement in a given formal language, and refines those probabilities over time. For instance, if the language is Peano arithmetic, it assigns…

Sentence

Increasing signal amplitude in electrical impedance tomography of neural activity using a parallel resistor inductor capacitor (RLC) circuit

2019-06-27

Objective: To increase the impedance signal amplitude produced during neural activity using a novel approach of implementing a parallel resistor inductor capacitor (RLC) circuit across the current source used in electric…

Design and Simulation of a Micro-coiled Digitally-Controlled Variable Inductor with a Monolithically Integrated MEMS Switch

2022-12-17 · Abdelhameed Sharaf, S. M. Eladl, A. Nasr, Mohamed Serry

This work introduces the design analysis simulation and a standard MEMS fabrication process for a three dimensional microcoil with a magnetic core and a digital switch configuration using a completely integrated fully ME…