paper-with-me

Papers

ActionFlowNet: Learning Motion Representation for Action Recognition

2016-12-09 · Joe Yue-Hei Ng, Jonghyun Choi, Jan Neumann, Larry S. Davis

Even with the recent advances in convolutional neural networks (CNN) in various visual recognition tasks, the state-of-the-art action recognition system still relies on hand crafted motion feature such as optical flow to achieve the best performance. We propose a multitask learning model ActionFlowNet to train a single stream network directly from raw pixels to jointly estimate optical flow while recognizing actions with convolutional neural networks, capturing both appearance and motion in a single model. We additionally provide insights to how the quality of the learned optical flow affects the action recognition. Our model significantly improves action recognition accuracy by a large margin 31% compared to state-of-the-art CNN-based action recognition models trained without external large scale data and additional optical flow input. Without pretraining on large external labeled datasets, our model, by well exploiting the motion information, achieves competitive recognition accuracy to the models trained with large labeled datasets such as ImageNet and Sport-1M.

📄 PDF Abstract BibTeX arXiv:1612.03052

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionOptical Flow EstimationTemporal Action Localization

Similar Papers 제목 키워드 기반

Learning multimodal representations for sample-efficient recognition of human actions

2019-03-06 · Miguel Vasco, Francisco S. Melo, David Martins de Matos, Ana Paiva 외

Humans interact in rich and diverse ways with the environment. However, the representation of such behavior by artificial agents is often limited. In this work we present \textit{motion concepts}, a novel multimodal repr…

MaCLR: Motion-aware Contrastive Learning of Representations for Videos

2021-06-17 · Fanyi Xiao, Joseph Tighe, Davide Modolo

We present MaCLR, a novel method to explicitly perform cross-modal self-supervised video representations learning from visual and motion modalities. Compared to previous video representation learning methods that mostly …

Action DetectionAction RecognitionContrastive LearningRepresentation Learning

Using phase instead of optical flow for action recognition

2018-09-10 · Omar Hommos, Silvia L. Pintea, Pascal S. M. Mettes, Jan C. van Gemert

Currently, the most common motion representation for action recognition is optical flow. Optical flow is based on particle tracking which adheres to a Lagrangian perspective on dynamics. In contrast to the Lagrangian per…

Action RecognitionMotion MagnificationOptical Flow EstimationTemporal Action Localization+1

Learning Joint Representation of Human Motion and Language

2022-10-27 · Jihoon Kim, Youngjae Yu, Seungyoun Shin, Taehyun Byun 외

In this work, we present MoLang (a Motion-Language connecting model) for learning joint representation of human motion and language, leveraging both unpaired and paired datasets of motion and language modalities. To this…

Action RecognitionContrastive LearningLanguage ModelingLanguage Modelling+1

One-shot action recognition in challenging therapy scenarios

2021-02-17 · Alberto Sabater, Laura Santos, Jose Santos-Victor, Alexandre Bernardino 외

One-shot action recognition aims to recognize new action categories from a single reference example, typically referred to as the anchor example. This work presents a novel approach for one-shot action recognition in the…

Action RecognitionOne-Shot 3D Action Recognition