paper-with-me

홈 › Papers

Learning Video Object Segmentation with Visual Memory

2017-04-19 · ICCV 2017 10 · Pavel Tokmakov, Karteek Alahari, Cordelia Schmid

This paper addresses the task of segmenting moving objects in unconstrained videos. We introduce a novel two-stream neural network with an explicit memory module to achieve this. The two streams of the network encode spatial and temporal features in a video sequence respectively, while the memory module captures the evolution of objects over time. The module to build a "visual memory" in video, i.e., a joint representation of all the video frames, is realized with a convolutional recurrent unit learned from a small number of training video sequences. Given a video frame as input, our approach assigns each pixel an object or background label based on the learned spatio-temporal features as well as the "visual memory" specific to the video, acquired automatically without any manually-annotated frames. The visual memory is implemented with convolutional gated recurrent units, which allows to propagate spatial information over time. We evaluate our method extensively on two benchmarks, DAVIS and Freiburg-Berkeley motion segmentation datasets, and show state-of-the-art results. For example, our approach outperforms the top method on the DAVIS dataset by nearly 6%. We also provide an extensive ablative analysis to investigate the influence of each component in the proposed framework.

📄 PDF Abstract BibTeX arXiv:1704.05737

Code (0)

등록된 구현이 없습니다.

Tasks

Motion SegmentationObjectSemantic SegmentationUnsupervised Video Object SegmentationVideo Object SegmentationVideo Semantic Segmentation

Similar Papers 제목 키워드 기반

Video Object Segmentation with Episodic Graph Memory Networks

2020-07-14 · ECCV 2020 8 · Xiankai Lu, Wenguan Wang, Martin Danelljan, Tianfei Zhou 외

How to make a segmentation model efficiently adapt to a specific video and to online target appearance variations are fundamentally crucial issues in the field of video object segmentation. In this work, a graph memory n…

ObjectSegmentationSemantic SegmentationVideo Object Segmentation+2

LIP: Learning Instance Propagation for Video Object Segmentation

2019-09-30 · Ye Lyu, George Vosselman, Gui-Song Xia, Michael Ying Yang

In recent years, the task of segmenting foreground objects from background in a video, i.e. video object segmentation (VOS), has received considerable attention. In this paper, we propose a single end-to-end trainable de…

Data AugmentationInstance SegmentationObjectSegmentation+3

Local-Global Context Aware Transformer for Language-Guided Video Segmentation

2022-03-18 · Chen Liang, Wenguan Wang, Tianfei Zhou, Jiaxu Miao 외

We explore the task of language-guided video segmentation (LVS). Previous algorithms mostly adopt 3D CNNs to learn video representation, struggling to capture long-term context and easily suffering from visual-linguistic…

Referring Expression SegmentationReferring Video Object SegmentationSegmentationSemantic Segmentation+4

Fast SAM2 with Text-Driven Token Pruning

2025-12-24 · Avilasha Mandal, Chaoning Zhang, Fachrina Dewi Puspitasari, Xudong Wang 외 arxiv

Segment Anything Model 2 (SAM2), a vision foundation model has significantly advanced in prompt-driven video object segmentation, yet their practical deployment remains limited by the high computational and memory cost o…

Video Object SegmentationVideo Segmentation

Learning to Segment Moving Objects

2017-12-01 · Pavel Tokmakov, Cordelia Schmid, Karteek Alahari

We study the problem of segmenting moving objects in unconstrained videos. Given a video, the task is to segment all the objects that exhibit independent motion in at least one frame. We formulate this as a learning prob…

Motion EstimationMotion SegmentationObjectObject Recognition+2