Human Action Recognition using Local Two-Stream Convolution Neural Network Features and Support Vector Machines
This paper proposes a simple yet effective method for human action recognition in video. The proposed method separately extracts local appearance and motion features using state-of-the-art three-dimensional convolutional neural networks from sampled snippets of a video. These local features are then concatenated to form global representations which are then used to train a linear SVM to perform the action classification using full context of the video, as partial context as used in previous works. The videos undergo two simple proposed preprocessing techniques, optical flow scaling and crop filling. We perform an extensive evaluation on three common benchmark dataset to empirically show the benefit of the SVM, and the two preprocessing steps.
Code (0)
등록된 구현이 없습니다.
Tasks
Action ClassificationAction RecognitionOptical Flow EstimationTemporal Action LocalizationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Global Temporal Representation based CNNs for Infrared Action Recognition
Infrared human action recognition has many advantages, i.e., it is insensitive to illumination change, appearance variability, and shadows. Existing methods for infrared action recognition are either based on spatial or …
Action RecognitionOptical Flow EstimationTemporal Action LocalizationA Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition
Human Interaction Recognition is the process of identifying interactive actions between multiple participants in a specific situation. The aim is to recognise the action interactions between multiple entities and their m…
Human Interaction RecognitionSpecificityMulti-View Region Adaptive Multi-temporal DMM and RGB Action Recognition
Human action recognition remains an important yet challenging task. This work proposes a novel action recognition system. It uses a novel Multiple View Region Adaptive Multi-resolution in time Depth Motion Map (MV-RAMDMM…
Action RecognitionHuman-Object Interaction DetectionTemporal Action LocalizationiPay: Integrated Payment Action Recognition via Multimodal Networks and Adaptive Spatial Prior Learning
Automated transit payment analysis is vital for scalable fare auditing and passenger analytics, yet practice still relies on limited manual inspection. Prior vision- and skeleton-based methods remain brittle under noisy …
Computational EfficiencyAction RecognitionTwo-stream Flow-guided Convolutional Attention Networks for Action Recognition
This paper proposes a two-stream flow-guided convolutional attention networks for action recognition in videos. The central idea is that optical flows, when properly compensated for the camera motion, can be used to guid…
Action RecognitionAction Recognition In VideosTemporal Action LocalizationVocal Bursts Valence Prediction