Deep Spatio-Temporal Random Fields for Efficient Video Segmentation
In this work we introduce a time- and memory-efficient method for structured prediction that couples neuron decisions across both space at time. We show that we are able to perform exact and efficient inference on a densely connected spatio-temporal graph by capitalizing on recent advances on deep Gaussian Conditional Random Fields (GCRFs). Our method, called VideoGCRF is (a) efficient, (b) has a unique global minimum, and (c) can be trained end-to-end alongside contemporary deep networks for video understanding. We experiment with multiple connectivity patterns in the temporal domain, and present empirical improvements over strong baselines on the tasks of both semantic and instance segmentation of videos.
Code (0)
등록된 구현이 없습니다.
Tasks
Instance SegmentationSemantic SegmentationStructured PredictionVideo SegmentationVideo Semantic SegmentationVideo UnderstandingSimilar Papers 제목 키워드 기반
Learning to Segment Moving Objects in Videos
We segment moving objects in videos by ranking spatio-temporal segment proposals according to "moving objectness": how likely they are to contain a moving object. In each video frame, we compute segment proposals using m…
SegmentationVideo SegmentationVideo Semantic SegmentationVideo-based Bottleneck Detection utilizing Lagrangian Dynamics in Crowded Scenes
Avoiding bottleneck situations in crowds is critical for the safety and comfort of people at large events or in public transportation. Based on the work of Lagrangian motion analysis we propose a novel video-based bottle…
Optical Flow EstimationSTFCN: Spatio-Temporal FCN for Semantic Video Segmentation
This paper presents a novel method to involve both spatial and temporal features for semantic video segmentation. Current work on convolutional neural networks(CNNs) has shown that CNNs provide advanced spatial features …
SegmentationSemantic SegmentationVideo SegmentationVideo Semantic SegmentationA Multi-Person Video Dataset Annotation Method of Spatio-Temporally Actions
Spatio-temporal action detection is an important and challenging problem in video understanding. However, the application of the existing large-scale spatio-temporal action datasets in specific fields is limited, and the…
Action DetectionVideo UnderstandingVideo Salient Object Detection Using Spatiotemporal Deep Features
This paper presents a method for detecting salient objects in videos where temporal information in addition to spatial information is fully taken into account. Following recent reports on the advantage of deep features o…
Objectobject-detectionObject DetectionRGB Salient Object Detection+5