First Person Action Recognition Using Deep Learned Descriptors
We focus on the problem of wearer's action recognition in first person a.k.a. egocentric videos. This problem is more challenging than third person activity recognition due to unavailability of wearer's pose and sharp movements in the videos caused by the natural head motion of the wearer. Carefully crafted features based on hands and objects cues for the problem have been shown to be successful for limited targeted datasets. We propose convolutional neural networks (CNNs) for end to end learning and classification of wearer's actions. The proposed network makes use of egocentric cues by capturing hand pose, head motion and saliency map. It is compact. It can also be trained from relatively small number of labeled egocentric videos that are available. We show that the proposed network can generalize and give state of the art performance on various disparate egocentric action datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionActivity RecognitionTemporal Action LocalizationSimilar Papers 제목 키워드 기반
Being the center of attention: A Person-Context CNN framework for Personality Recognition
This paper proposes a novel study on personality recognition using video data from different scenarios. Our goal is to jointly model nonverbal behavioral cues with contextual information for a robust, multi-scenario, per…
Boosted Multiple Kernel Learning for First-Person Activity Recognition
Activity recognition from first-person (ego-centric) videos has recently gained attention due to the increasing ubiquity of the wearable cameras. There has been a surge of efforts adapting existing feature descriptors an…
Activity RecognitionPooled Motion Features for First-Person Videos
In this paper, we present a new feature representation for first-person videos. In first-person video understanding (e.g., activity recognition), it is very important to capture both entire scene dynamics (i.e., egomotio…
Activity RecognitionActivity Recognition In VideosTime SeriesTime Series Analysis+1Revisiting Human Action Recognition: Personalization vs. Generalization
By thoroughly revisiting the classic human action recognition paradigm, this paper aims at proposing a new approach for the design of effective action classification systems. Taking as testbed publicly available three-di…
Action ClassificationAction RecognitionGeneral ClassificationTemporal Action LocalizationEXMOVES: Classifier-based Features for Scalable Action Recognition
This paper introduces EXMOVES, learned exemplar-based features for efficient recognition of actions in videos. The entries in our descriptor are produced by evaluating a set of movement classifiers over spatial-temporal …
Action RecognitionGeneral ClassificationTemporal Action Localization