paper-with-me

Papers

Representing Videos Using Mid-level Discriminative Patches

2013-06-01 · CVPR 2013 6 · Arpit Jain, Abhinav Gupta, Mikel Rodriguez, Larry S. Davis

representation for videos based on mid-level discriminative spatio-temporal patches. These spatio-temporal patches might correspond to a primitive human action, a semantic object, or perhaps a random but informative spatiotemporal patch in the video. What defines these spatiotemporal patches is their discriminative and representative properties. We automatically mine these patches from hundreds of training videos and experimentally demonstrate that these patches establish correspondence across videos and align the videos for label transfer techniques. Furthermore, these patches can be used as a discriminative vocabulary for action classification where they demonstrate stateof-the-art performance on UCF50 and Olympics datasets.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action ClassificationGeneral Classification

Similar Papers 제목 키워드 기반

Mid-level Representation for Visual Recognition

2015-12-23 · Moin Nabi

Visual Recognition is one of the fundamental challenges in AI, where the goal is to understand the semantics of visual data. Employing mid-level representation, in particular, shifted the paradigm in visual recognition. …

object-detectionObject DetectionVideo Understanding

Harvesting Discriminative Meta Objects with Deep CNN Features for Scene Classification

2015-10-06 · ICCV 2015 12 · Ruobing Wu, Baoyuan Wang, Wenping Wang, Yizhou Yu

Recent work on scene classification still makes use of generic CNN features in a rudimentary manner. In this ICCV 2015 paper, we present a novel pipeline built upon deep CNN features to harvest discriminative visual obje…

ClusteringGeneral ClassificationRegion ProposalScene Classification+1

Real-Time Anomalous Behavior Detection and Localization in Crowded Scenes

2015-11-21 · Mohammad Sabokrou, Mahmood Fathy, Mojtaba Hosseini

In this paper, we propose an accurate and real-time anomaly detection and localization in crowded scenes, and two descriptors for representing anomalous behavior in video are proposed. We consider a video as being a set …

Anomaly Detection

Patch-Based Discriminative Feature Learning for Unsupervised Person Re-Identification

2019-06-01 · CVPR 2019 6 · Qize Yang, Hong-Xing Yu, Ancong Wu, Wei-Shi Zheng

While discriminative local features have been shown effective in solving the person re-identification problem, they are limited to be trained on fully pairwise labelled data which is expensive to obtain. In this work, we…

Person Re-IdentificationUnsupervised Person Re-Identification

Action Recognition by Hierarchical Mid-level Action Elements

2015-08-31 · ICCV 2015 12 · Tian Lan, Yuke Zhu, Amir Roshan Zamir, Silvio Savarese

Realistic videos of human actions exhibit rich spatiotemporal structures at multiple levels of granularity: an action can always be decomposed into multiple finer-grained elements in both space and time. To capture this …

Action ParsingAction RecognitionClusteringTemporal Action Localization