paper-with-me

홈 › Papers

Multi-Cue Structure Preserving MRF for Unconstrained Video Segmentation

2015-06-30 · ICCV 2015 12 · Saehoon Yi, Vladimir Pavlovic

Video segmentation is a stepping stone to understanding video context. Video segmentation enables one to represent a video by decomposing it into coherent regions which comprise whole or parts of objects. However, the challenge originates from the fact that most of the video segmentation algorithms are based on unsupervised learning due to expensive cost of pixelwise video annotation and intra-class variability within similar unconstrained video classes. We propose a Markov Random Field model for unconstrained video segmentation that relies on tight integration of multiple cues: vertices are defined from contour based superpixels, unary potentials from temporal smooth label likelihood and pairwise potentials from global structure of a video. Multi-cue structure is a breakthrough to extracting coherent object regions for unconstrained videos in absence of supervision. Our experiments on VSB100 dataset show that the proposed model significantly outperforms competing state-of-the-art algorithms. Qualitative analysis illustrates that video segmentation result of the proposed model is consistent with human perception of objects.

📄 PDF Abstract BibTeX arXiv:1506.09124

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSuperpixelsVideo SegmentationVideo Semantic Segmentation

Similar Papers 제목 키워드 기반

Motion-Appearance Interactive Encoding for Object Segmentation in Unconstrained Videos

2017-07-25 · Chunchao Guo, Jian-Huang Lai, Xiaohua Xie

We present a novel method of integrating motion and appearance cues for foreground object segmentation in unconstrained videos. Unlike conventional methods encoding motion and appearance patterns individually, our method…

Graph MatchingObjectObject LocalizationSemantic Segmentation+1

3PoinTr: 3D Point Tracks for Learning Manipulation from Unconstrained Human Videos

2026-03-09 · Adam Hung, Bardienus Pieter Duisterhof, Jeffrey Ichnowski arxiv

Learning manipulation policies from human videos could greatly reduce the need for expensive robot demonstrations, but existing approaches typically require restrictive assumptions such as choreographed human motions, pr…

WildActor: Unconstrained Identity-Preserving Video Generation

2026-02-28 · Qin Guo, Tianyu Yang, Xuanhua He, Fei Shen 외 arxiv

Production-ready human video generation requires digital actors to maintain strictly consistent full-body identities across dynamic shots, viewpoints and motions, a setting that remains challenging for existing methods. …

Video Generation

Towards Automatic Learning of Procedures from Web Instructional Videos

2017-03-28 · Luowei Zhou, Chenliang Xu, Jason J. Corso

The potential for agents, whether embodied or software, to learn by observing other agents performing procedures involving objects and actions is rich. Current research on automatic procedure learning heavily relies on a…

Dense Video CaptioningProcedure LearningSegmentationVideo Captioning

Unconstrained Monocular 3D Human Pose Estimation by Action Detection and Cross-Modality Regression Forest

2013-06-01 · CVPR 2013 6 · Tsz-Ho Yu, Tae-Kyun Kim, Roberto Cipolla

This work addresses the challenging problem of unconstrained 3D human pose estimation (HPE) from a novel perspective. Existing approaches struggle to operate in realistic applications, mainly due to their scene-dependent…

2D Pose Estimation3D Human Pose EstimationAction DetectionMonocular 3D Human Pose Estimation+3