Curvature: A signature for Action Recognition in Video Sequences
In this paper, a novel signature of human action recognition, namely the curvature of a video sequence, is introduced. In this way, the distribution of sequential data is modeled, which enables few-shot learning. Instead of depending on recognizing features within images, our algorithm views actions as sequences on the universal time scale across a whole sequence of images. The video sequence, viewed as a curve in pixel space, is aligned by reparameterization using the arclength of the curve in pixel space. Once such curvatures are obtained, statistical indexes are extracted and fed into a learning-based classifier. Overall, our method is simple but powerful. Preliminary experimental results show that our method is effective and achieves state-of-the-art performance in video-based human action recognition. Moreover, we see latent capacity in transferring this idea into other sequence-based recognition applications such as speech recognition, machine translation, and text generation.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionFew-Shot LearningMachine Translationspeech-recognitionSpeech RecognitionTemporal Action LocalizationText GenerationTranslationSimilar Papers 제목 키워드 기반
DASZL: Dynamic Action Signatures for Zero-shot Learning
There are many realistic applications of activity recognition where the set of potential activity descriptions is combinatorially large. This makes end-to-end supervised training of a recognition system impractical as no…
Action DetectionActivity DetectionActivity RecognitionVideo Classification+1Egocentric View Hand Action Recognition by Leveraging Hand Surface and Hand Grasp Type
We introduce a multi-stage framework that uses mean curvature on a hand surface and focuses on learning interaction between hand and object by analyzing hand grasp type for hand action recognition in egocentric videos. T…
Action RecognitionObjectVocal Bursts Type PredictionRNN Fisher Vectors for Action Recognition and Image Annotation
Recurrent Neural Networks (RNNs) have had considerable success in classifying and predicting sequences. We demonstrate that RNNs can be effectively used in order to encode sequences and provide effective representations.…
Action RecognitionTemporal Action LocalizationTransfer LearningArabic Text Recognition in Video Sequences
In this paper, we propose a robust approach for text extraction and recognition from Arabic news video sequence. The text included in video sequences is an important needful for indexing and searching system. However, th…
Emotion recognition in talking-face videos using persistent entropy and neural networks
The automatic recognition of a person's emotional state has become a very active research field that involves scientists specialized in different areas such as artificial intelligence, computer vision or psychology, amon…
Emotion Recognition