paper-with-me

Papers

Multi-Source Fusion and Automatic Predictor Selection for Zero-Shot Video Object Segmentation

2021-08-11 · Xiaoqi Zhao, Youwei Pang, Jiaxing Yang, Lihe Zhang, Huchuan Lu

Location and appearance are the key cues for video object segmentation. Many sources such as RGB, depth, optical flow and static saliency can provide useful information about the objects. However, existing approaches only utilize the RGB or RGB and optical flow. In this paper, we propose a novel multi-source fusion network for zero-shot video object segmentation. With the help of interoceptive spatial attention module (ISAM), spatial importance of each source is highlighted. Furthermore, we design a feature purification module (FPM) to filter the inter-source incompatible features. By the ISAM and FPM, the multi-source features are effectively fused. In addition, we put forward an automatic predictor selection network (APS) to select the better prediction of either the static saliency predictor or the moving object predictor in order to prevent over-reliance on the failed results caused by low-quality optical flow maps. Extensive experiments on three challenging public benchmarks (i.e. DAVIS$_{16}$, Youtube-Objects and FBMS) show that the proposed model achieves compelling performance against the state-of-the-arts. The source code will be publicly available at \textcolor{red}{\url{https://github.com/Xiaoqi-Zhao-DLUT/Multi-Source-APS-ZVOS}}.

📄 PDF Abstract BibTeX arXiv:2108.05076

Code (1)

xiaoqi-zhao-dlut/multi-source-aps-zvos 공식 구현 pytorch

Tasks

Depth EstimationObjectSalient Object DetectionUnsupervised Video Object SegmentationVideo Object SegmentationZero-Shot Video Object Segmentation

Similar Papers 제목 키워드 기반

Diffusion-Driven High-Dimensional Variable Selection

2025-08-19 · Minjie Wang, Xiaotong Shen, Wei Pan arxiv

Variable selection for high-dimensional, highly correlated data has long been a challenging problem, often yielding unstable and unreliable models. We propose a resample-aggregate framework that exploits diffusion models…

Transfer LearningData Augmentation

Controllable Accent Normalization via Discrete Diffusion

2026-03-15 · Qibing Bai, Yuhan Du, Tom Ko, Shuai Wang 외 arxiv

Existing accent normalization methods do not typically offer control over accent strength, yet many applications-such as language learning and dubbing-require tunable accent retention. We propose DLM-AN, a controllable a…

Adaptive Multi-source Predictor for Zero-shot Video Object Segmentation

2023-03-18 · Xiaoqi Zhao, Shijie Chang, Youwei Pang, Jiaxing Yang 외

Static and moving objects often occur in real-life videos. Most video object segmentation methods only focus on extracting and exploiting motion cues to perceive moving objects. Once faced with the frames of static objec…

ObjectOptical Flow EstimationSemantic SegmentationUnsupervised Video Object Segmentation+3

Joint Manifold Diffusion for Combining Predictions on Decoupled Observations

2019-04-10 · CVPR 2019 6 · Kwang In Kim, Hyung Jin Chang

We present a new predictor combination algorithm that improves a given task predictor based on potentially relevant reference predictors. Existing approaches are limited in that, to discover the underlying task dependenc…

Source Data Selection for Brain-Computer Interfaces based on Simple Features

2024-10-03 · Frida Heskebeck, Carolina Bergeling, Bo Bernhardsson

This paper demonstrates that simple features available during the calibration of a brain-computer interface can be utilized for source data selection to improve the performance of the brain-computer interface for a new t…

Brain Computer InterfaceMotor ImageryTransfer Learning