paper-with-me

홈 › Papers

Video Event Recognition With Deep Hierarchical Context Model

2015-06-01 · CVPR 2015 6 · Xiaoyang Wang, Qiang Ji

Video event recognition still faces great challenges due to large intra-class variation and low image resolution, in particular for surveillance videos. To mitigate these challenges and to improve the event recognition performance, various context information from the feature level, the semantic level, as well as the prior level is utilized. Different from most existing context approaches that utilize context in one of the three levels through shallow models like support vector machines, or probabilistic models like BN and MRF, we propose a deep hierarchical context model that simultaneously learns and integrates context at all three levels, and holistically utilizes the integrated contexts for event recognition. We first introduce two types of context features describing the event neighborhood, and then utilize the proposed deep model to learn the middle level representations and combine the bottom feature level, middle semantic level and top prior level contexts together for event recognition. The experiments on state of art surveillance video event benchmarks including VIRAT 1.0 Ground Dataset, VIRAT 2.0 Ground Dataset, and the UT-Interaction Dataset demonstrate that the proposed model is quite effective in utilizing the context information for event recognition. It outperforms the existing context approaches that also utilize multiple level contexts on these event benchmarks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

model

Similar Papers 제목 키워드 기반

A Hierarchical Context Model for Event Recognition in Surveillance Video

2014-06-27 · The IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2014 2014 6 · Xiaoyang Wang, Qiang Ji

Due to great challenges such as tremendous intra-class variations and low image resolution, context information has been playing a more and more important role for accurate and robust event recognition in surveillance vi…

Action Recognition

Hierarchical Self-supervised Representation Learning for Movie Understanding

2022-04-06 · CVPR 2022 1 · Fanyi Xiao, Kaustav Kundu, Joseph Tighe, Davide Modolo

Most self-supervised video representation learning approaches focus on action recognition. In contrast, in this paper we focus on self-supervised video learning for movie understanding and propose a novel hierarchical se…

Action RecognitionContrastive LearningRepresentation Learning

Hierarchical Context-aware Network for Dense Video Event Captioning

2021-08-01 · ACL 2021 5 · Lei Ji, Xianglin Guo, Haoyang Huang, Xilin Chen

Dense video event captioning aims to generate a sequence of descriptive captions for each event in a long untrimmed video. Video-level context provides important information and facilities the model to generate consisten…

Descriptive

Hierarchical Object-oriented Spatio-Temporal Reasoning for Video Question Answering

2021-06-25 · Long Hoang Dang, Thao Minh Le, Vuong Le, Truyen Tran

Video Question Answering (Video QA) is a powerful testbed to develop new AI capabilities. This task necessitates learning to reason about objects, relations, and events across visual and linguistic domains in space-time.…

ObjectQuestion AnsweringVideo Question Answering

Hierarchical Video Frame Sequence Representation with Deep Convolutional Graph Network

2019-06-02 · Feng Mao, Xiang Wu, Hui Xue, Rong Zhang

High accuracy video label prediction (classification) models are attributed to large scale data. These data could be frame feature sequences extracted by a pre-trained convolutional-neural-network, which promote the effi…

General ClassificationGraph Neural NetworkVideo ClassificationVideo Understanding