Learning Dynamic Hierarchical Models for Anytime Scene Labeling
With increasing demand for efficient image and video analysis, test-time cost of scene parsing becomes critical for many large-scale or time-sensitive vision applications. We propose a dynamic hierarchical model for anytime scene labeling that allows us to achieve flexible trade-offs between efficiency and accuracy in pixel-level prediction. In particular, our approach incorporates the cost of feature computation and model inference, and optimizes the model performance for any given test-time budget by learning a sequence of image-adaptive hierarchical models. We formulate this anytime representation learning as a Markov Decision Process with a discrete-continuous state-action space. A high-quality policy of feature and model selection is learned based on an approximate policy iteration method with action proposal mechanism. We demonstrate the advantages of our dynamic non-myopic anytime scene parsing on three semantic segmentation datasets, which achieves $90\%$ of the state-of-the-art performances by using $15\%$ of their overall costs.
Code (0)
등록된 구현이 없습니다.
Tasks
Model SelectionRepresentation LearningScene LabelingScene ParsingSemantic SegmentationSimilar Papers 제목 키워드 기반
Anytime Recognition of Objects and Scenes
Humans are capable of perceiving a scene at a glance, and obtain deeper understanding with additional time. Similarly, visual recognition deployments should be robust to varying computational budgets. Such situations req…
General ClassificationObject RecognitionTowards Anytime Retrieval: A Benchmark for Anytime Person Re-Identification
In real applications, person re-identification (ReID) is expected to retrieve the target person at any time, including both daytime and nighttime, ranging from short-term to long-term. However, existing ReID tasks and da…
Person Re-IdentificationScene Labeling with Contextual Hierarchical Models
Scene labeling is the problem of assigning an object label to each pixel. It unifies the image segmentation and object recognition problems. The importance of using contextual information in scene labeling frameworks has…
Edge DetectionImage SegmentationObjectObject Recognition+3Anytime Hierarchical Clustering
We propose a new anytime hierarchical clustering method that iteratively transforms an arbitrary initial hierarchy on the configuration of measurements along a sequence of trees we prove for a fixed data set must termina…
Anomaly DetectionClusteringGeometric Scene Parsing with Hierarchical LSTM
This paper addresses the problem of geometric scene parsing, i.e. simultaneously labeling geometric surfaces (e.g. sky, ground and vertical plane) and determining the interaction relations (e.g. layering, supporting, sid…
3D ReconstructionRelation PredictionScene LabelingScene Parsing