DIFEM: Key-points Interaction based Feature Extraction Module for Violence Recognition in Videos
Violence detection in surveillance videos is a critical task for ensuring public safety. As a result, there is increasing need for efficient and lightweight systems for automatic detection of violent behaviours. In this work, we propose an effective method which leverages human skeleton key-points to capture inherent properties of violence, such as rapid movement of specific joints and their close proximity. At the heart of our method is our novel Dynamic Interaction Feature Extraction Module (DIFEM) which captures features such as velocity, and joint intersections, effectively capturing the dynamics of violent behavior. With the features extracted by our DIFEM, we use various classification algorithms such as Random Forest, Decision tree, AdaBoost and k-Nearest Neighbor. Our approach has substantially lesser amount of parameter expense than the existing state-of-the-art (SOTA) methods employing deep learning techniques. We perform extensive experiments on three standard violence recognition datasets, showing promising performance in all three datasets. Our proposed method surpasses several SOTA violence recognition methods.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Effective Action Recognition with Embedded Key Point Shifts
Temporal feature extraction is an essential technique in video-based action recognition. Key points have been utilized in skeleton-based action recognition methods but they require costly key point annotation. In this pa…
Action RecognitionSkeleton Based Action RecognitionHOKEM: Human and Object Keypoint-based Extension Module for Human-Object Interaction Detection
Human-object interaction (HOI) detection for capturing relationships between humans and objects is an important task in the semantic understanding of images. When processing human and object keypoints extracted from an i…
Human-Object Interaction DetectionObjectPointSCNet: Point Cloud Structure and Correlation Learning Based on Space Filling Curve-Guided Sampling
Geometrical structures and the internal local region relationship, such as symmetry, regular array, junction, etc., are essential for understanding a 3D shape. This paper proposes a point cloud feature extraction network…
3D Point Cloud ClassificationSemantic SegmentationEmotion-cause pair extraction method based on multi-granularity information and multi-module interaction
The purpose of emotion-cause pair extraction is to extract the pair of emotion clauses and cause clauses. On the one hand, the existing methods do not take fully into account the relationship between the emotion extracti…
Emotion-Cause Pair ExtractionPositionAppformer: A Novel Framework for Mobile App Usage Prediction Leveraging Progressive Multi-Modal Data Fusion and Feature Extraction
This article presents Appformer, a novel mobile application prediction framework inspired by the efficiency of Transformer-like architectures in processing sequential data through self-attention mechanisms. Combining a M…
Time Series AnalysisWord Embeddings