paper-with-me

Papers

Graph-Jigsaw Conditioned Diffusion Model for Skeleton-based Video Anomaly Detection

2024-03-18 · Ali Karami, Thi Kieu Khanh Ho, Narges Armanfard

Skeleton-based video anomaly detection (SVAD) is a crucial task in computer vision. Accurately identifying abnormal patterns or events enables operators to promptly detect suspicious activities, thereby enhancing safety. Achieving this demands a comprehensive understanding of human motions, both at body and region levels, while also accounting for the wide variations of performing a single action. However, existing studies fail to simultaneously address these crucial properties. This paper introduces a novel, practical and lightweight framework, namely Graph-Jigsaw Conditioned Diffusion Model for Skeleton-based Video Anomaly Detection (GiCiSAD) to overcome the challenges associated with SVAD. GiCiSAD consists of three novel modules: the Graph Attention-based Forecasting module to capture the spatio-temporal dependencies inherent in the data, the Graph-level Jigsaw Puzzle Maker module to distinguish subtle region-level discrepancies between normal and abnormal motions, and the Graph-based Conditional Diffusion model to generate a wide spectrum of human motions. Extensive experiments on four widely used skeleton-based video datasets show that GiCiSAD outperforms existing methods with significantly fewer training parameters, establishing it as the new state-of-the-art.

📄 PDF Abstract BibTeX arXiv:2403.12172

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionGraph AttentionVideo Anomaly Detection

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Jigsaw Jigsaw is a self-supervision approach that relies on jigsaw-like puzzles as the pretext task in order to learn image representations.

Similar Papers 제목 키워드 기반

Solving Masked Jigsaw Puzzles with Diffusion Vision Transformers

2024-04-10 · CVPR 2024 1 · Jinyang Liu, Wondmgezahu Teshome, Sandesh Ghimire, Mario Sznaier 외

Solving image and video jigsaw puzzles poses the challenging task of rearranging image fragments or video frames from unordered sequences to restore meaningful images and video sequences. Existing approaches often hinge …

Democratizing High-Fidelity Co-Speech Gesture Video Generation

2025-07-09 · Xu Yang, Shaoli Huang, Shenbo Xie, Xuelin Chen 외 arxiv

Co-speech gesture video generation aims to synthesize realistic, audio-aligned videos of speakers, complete with synchronized facial expressions and body gestures. This task presents challenges due to the significant one…

Video Generation

GALA: Graph Diffusion-based Alignment with Jigsaw for Source-free Domain Adaptation

2024-10-22 · Junyu Luo, Yiyang Gu, Xiao Luo, Wei Ju 외

Source-free domain adaptation is a crucial machine learning topic, as it contains numerous applications in the real world, particularly with respect to data privacy. Existing approaches predominantly focus on Euclidean d…

Domain AdaptationGRAPH DOMAIN ADAPTATIONGraph Neural NetworkSource-Free Domain Adaptation

Video Motion Graphs

2025-03-26 · Haiyang Liu, Zhan Xu, Fa-Ting Hong, Hsin-Ping Huang 외

We present Video Motion Graphs, a system designed to generate realistic human motion videos. Using a reference video and conditional signals such as music or motion tags, the system synthesizes new videos by first retrie…

Motion InterpolationVideo Frame InterpolationVideo Generation

Controllable Complex Human Motion Video Generation via Text-to-Skeleton Cascades

2026-03-09 · Ashkan Taghipour, Morteza Ghahremani, Zinuo Li, Hamid Laga 외 arxiv

Generating videos of complex human motions such as flips, cartwheels, and martial arts remains challenging for current video diffusion models. Text-only conditioning is temporally ambiguous for fine-grained motion contro…

Video Generation