paper-with-me

Papers

Motion-Guided Cascaded Refinement Network for Video Object Segmentation

2018-06-01 · CVPR 2018 6 · Ping Hu, Gang Wang, Xiangfei Kong, Jason Kuen, Yap-Peng Tan

Deep CNNs have achieved superior performance in many tasks of computer vision and image understanding. However, it is still difficult to effectively apply deep CNNs to video object segmentation(VOS) since treating video frames as separate and static will lose the information hidden in motion. To tackle this problem, we propose a Motion-guided Cascaded Refinement Network for VOS. By assuming the object motion is normally different from the background motion, for a video frame we first apply an active contour model on optical flow to coarsely segment objects of interest. Then, the proposed Cascaded Refinement Network(CRN) takes the coarse segmentation as guidance to generate an accurate segmentation of full resolution. In this way, the motion information and the deep CNNs can well complement each other to accurately segment objects from video frames. Furthermore, in CRN we introduce a Single-channel Residual Attention Module to incorporate the coarse segmentation map as attention, making our network effective and efficient in both training and testing. We perform experiments on the popular benchmarks and the results show that our method achieves state-of-the-art performance at a much faster speed.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectOptical Flow EstimationSegmentationSemantic SegmentationVideo Object SegmentationVideo Semantic Segmentation

Similar Papers 제목 키워드 기반

2nd Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation

2024-06-20 · Bin Cao, Yisi Zhang, Xuanxu Lin, Xingjian He 외

Motion Expression guided Video Segmentation is a challenging task that aims at segmenting objects in the video based on natural language expressions with motion descriptions. Unlike the previous referring video object se…

Instance SegmentationReferring Video Object SegmentationSegmentationSemantic Segmentation+4

STSeg-Complex Video Object Segmentation: The 1st Solution for 4th PVUW MOSE Challenge

2025-04-11 · Kehuan Song, Xinglin Xie, Kexin Zhang, Licheng Jiao 외

Segmentation of video objects in complex scenarios is highly challenging, and the MOSE dataset has significantly contributed to the development of this field. This technical report details the STSeg solution proposed by …

Semantic SegmentationVideo Object SegmentationVideo Semantic Segmentation

Visually Guided Sound Source Separation using Cascaded Opponent Filter Network

2020-06-04 · Lingyu Zhu, Esa Rahtu

The objective of this paper is to recover the original component signals from a mixture audio with the aid of visual cues of the sound sources. Such task is usually referred as visually guided sound source separation. Th…

Visually Guided Sound Source Separation

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts

2026-02-12 · Chen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang 외 arxiv

Recent advances in video diffusion models have significantly improved visual quality, yet ultra-high-resolution (UHR) video generation remains a formidable challenge due to the compounded difficulties of motion modeling,…

Video Generation

Motion-Adaptive Temporal Attention for Lightweight Video Generation with Stable Diffusion

2026-03-18 · Rui Hong, Shuxue Quan arxiv

We present a motion-adaptive temporal attention mechanism for parameter-efficient video generation built upon frozen Stable Diffusion models. Rather than treating all video content uniformly, our method dynamically adjus…

Video Generation