paper-with-me

Papers

Deep Spatio-temporal Manifold Network for Action Recognition

2017-05-09 · Ce Li, Chen Chen, Baochang Zhang, Qixiang Ye, Jungong Han, Rongrong Ji

Visual data such as videos are often sampled from complex manifold. We propose leveraging the manifold structure to constrain the deep action feature learning, thereby minimizing the intra-class variations in the feature space and alleviating the over-fitting problem. Considering that manifold can be transferred, layer by layer, from the data domain to the deep features, the manifold priori is posed from the top layer into the back propagation learning procedure of convolutional neural network (CNN). The resulting algorithm --Spatio-Temporal Manifold Network-- is solved with the efficient Alternating Direction Method of Multipliers and Backward Propagation (ADMM-BP). We theoretically show that STMN recasts the problem as projection over the manifold via an embedding method. The proposed approach is evaluated on two benchmark datasets, showing significant improvements to the baselines.

📄 PDF Abstract BibTeX arXiv:1705.03148

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionTemporal Action Localization

Similar Papers 제목 키워드 기반

Learning Expressionlets on Spatio-Temporal Manifold for Dynamic Facial Expression Recognition

2014-06-01 · CVPR 2014 6 · Mengyi Liu, Shiguang Shan, Ruiping Wang, Xilin Chen

Facial expression is temporally dynamic event which can be decomposed into a set of muscle motions occurring in different facial regions over various time intervals. For dynamic expression recognition, two key issues, te…

Dynamic Facial Expression RecognitionFacial Expression RecognitionFacial Expression Recognition (FER)

Log-Euclidean Bag of Words for Human Action Recognition

2014-06-09 · Masoud Faraki, Maziar Palhang, Conrad Sanderson

Representing videos by densely extracted local space-time features has recently become a popular approach for analysing actions. In this paper, we tackle the problem of categorising human actions by devising Bag of Words…

Action RecognitionOptical Flow EstimationTemporal Action Localization

Multi-stage Factorized Spatio-Temporal Representation for RGB-D Action and Gesture Recognition

2023-08-23 · Yujun Ma, Benjia Zhou, Ruili Wang, Pichao Wang

RGB-D action and gesture recognition remain an interesting topic in human-centered scene understanding, primarily due to the multiple granularities and large variation in human motion. Although many RGB-D based action an…

Gesture RecognitionScene Understanding

Spatiotemporal Filtering for Event-Based Action Recognition

2019-03-17 · Rohan Ghosh, Anupam Gupta, Andrei Nakagawa, Alcimar Soares 외

In this paper, we address the challenging problem of action recognition, using event-based cameras. To recognise most gestural actions, often higher temporal precision is required for sampling visual information. Actions…

Action RecognitionTemporal Action Localization

Spatiotemporal Residual Networks for Video Action Recognition

2016-11-07 · NeurIPS 2016 12 · Christoph Feichtenhofer, Axel Pinz, Richard P. Wildes

Two-stream Convolutional Networks (ConvNets) have shown strong performance for human action recognition in videos. Recently, Residual Networks (ResNets) have arisen as a new technique to train extremely deep architecture…

Action RecognitionAction Recognition In VideosTemporal Action Localization