paper-with-me

홈 › Papers

Towards Good Practices for Action Video Encoding

2014-06-01 · CVPR 2014 6 · Jianxin Wu, Yu Zhang, Weiyao Lin

High dimensional representations such as VLAD or FV have shown excellent accuracy in action recognition. This paper shows that a proper encoding built upon VLAD can achieve further accuracy boost with only negligible computational cost. We empirically evaluated various VLAD improvement technologies to determine good practices in VLAD-based video encoding. Furthermore, we propose an interpretation that VLAD is a maximum entropy linear feature learning process. Combining this new perspective with observed VLAD data distribution properties, we propose a simple, lightweight, but powerful bimodal encoding method. Evaluated on 3 benchmark action recognition datasets (UCF101, HMDB51 and Youtube), the bimodal encoding improves VLAD by large margins in action recognition.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionTemporal Action Localization

Similar Papers 제목 키워드 기반

Temporal Segment Networks: Towards Good Practices for Deep Action Recognition

2016-08-02 · Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao 외

Deep convolutional networks have achieved great success for visual recognition in still images. However, for action recognition in videos, the advantage over traditional methods is not so evident. This paper aims to disc…

Action ClassificationAction RecognitionAction Recognition In VideosMultimodal Activity Recognition+1

Temporal Segment Networks for Action Recognition in Videos

2017-05-08 · Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao 외

Deep convolutional networks have achieved great success for image recognition. However, for action recognition in videos, their advantage over traditional methods is not so evident. We present a general and flexible vide…

Action ClassificationAction RecognitionAction Recognition In VideosMicro-Action Recognition+2

Towards Good Practices for Very Deep Two-Stream ConvNets

2015-07-08 · Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao

Deep convolutional networks have achieved great success for object recognition in still images. However, for action recognition in videos, the improvement of deep convolutional networks is not so evident. We argue that t…

Action RecognitionAction Recognition In VideosComputational EfficiencyData Augmentation+3

Best Practices for 2-Body Pose Forecasting

2023-04-12 · Muhammad Rameez Ur Rahman, Luca Scofano, Edoardo De Matteis, Alessandro Flaborea 외

The task of collaborative human pose forecasting stands for predicting the future poses of multiple interacting people, given those in previous frames. Predicting two people in interaction, instead of each separately, pr…

Human Pose ForecastingMotion Forecastingmotion predictionMulti-Person Pose forecasting

Towards Good Practices for Video Object Segmentation

2019-09-30 · Dongdong Yu, Kai Su, Hengkai Guo, Jian Wang 외

Semi-supervised video object segmentation is an interesting yet challenging task in machine learning. In this work, we conduct a series of refinements with the propagation-based video object segmentation method and empir…

BIG-bench Machine LearningObjectOne-shot visual object segmentationSegmentation+4