The AVA-Kinetics Localized Human Actions Video Dataset
This paper describes the AVA-Kinetics localized human actions video dataset. The dataset is collected by annotating videos from the Kinetics-700 dataset using the AVA annotation protocol, and extending the original AVA dataset with these new AVA annotated Kinetics clips. The dataset contains over 230k clips annotated with the 80 AVA action classes for each of the humans in key-frames. We describe the annotation process and provide statistics about the new dataset. We also include a baseline evaluation using the Video Action Transformer Network on the AVA-Kinetics dataset, demonstrating improved performance for action classification on the AVA test set. The dataset can be downloaded from https://research.google.com/ava/
Code (0)
등록된 구현이 없습니다.
Tasks
Action ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Handcrafted localized phase features for human action recognition
Human action recognition is one of the most important topics in computer vision. Monitoring elderly people and children, smart surveillance systems and human-computer interaction are a few examples of its applications. T…
Action ClassificationAction RecognitionOptical Flow EstimationTemporal Action LocalizationThe Kinetics Human Action Video Dataset
We describe the DeepMind Kinetics human action video dataset. The dataset contains 400 human action classes, with at least 400 video clips for each action. Each clip lasts around 10s and is taken from a different YouTube…
Action ClassificationGeneral ClassificationI3D-LSTM: A New Model for Human Action Recognition
Action recognition has already been a heated research topic recently, which attempts to classify different human actions in videos. The current main-stream methods generally utilize ImageNet-pretrained model as features …
Action RecognitionTemporal Action LocalizationA Short Note on the Kinetics-700-2020 Human Action Dataset
We describe the 2020 edition of the DeepMind Kinetics human action dataset, which replenishes and extends the Kinetics-700 dataset. In this new version, there are at least 700 video clips from different YouTube videos fo…
Three Branches: Detecting Actions With Richer Features
We present our three branch solutions for International Challenge on Activity Recognition at CVPR2019. This model seeks to fuse richer information of global video clip, short human attention and long-term human activity …
Action LocalizationActivity RecognitionSpatio-Temporal Action LocalizationTemporal Action Localization