paper-with-me

Papers

Siamese Network for RGB-D Salient Object Detection and Beyond

2020-08-26 · Keren Fu, Deng-Ping Fan, Ge-Peng Ji, Qijun Zhao, Jianbing Shen, Ce Zhu

Existing RGB-D salient object detection (SOD) models usually treat RGB and depth as independent information and design separate networks for feature extraction from each. Such schemes can easily be constrained by a limited amount of training data or over-reliance on an elaborately designed training process. Inspired by the observation that RGB and depth modalities actually present certain commonality in distinguishing salient objects, a novel joint learning and densely cooperative fusion (JL-DCF) architecture is designed to learn from both RGB and depth inputs through a shared network backbone, known as the Siamese architecture. In this paper, we propose two effective components: joint learning (JL), and densely cooperative fusion (DCF). The JL module provides robust saliency feature learning by exploiting cross-modal commonality via a Siamese network, while the DCF module is introduced for complementary feature discovery. Comprehensive experiments using five popular metrics show that the designed framework yields a robust RGB-D saliency detector with good generalization. As a result, JL-DCF significantly advances the state-of-the-art models by an average of ~2.0% (max F-measure) across seven challenging datasets. In addition, we show that JL-DCF is readily applicable to other related multi-modal detection tasks, including RGB-T (thermal infrared) SOD and video SOD, achieving comparable or even better performance against state-of-the-art methods. We also link JL-DCF to the RGB-D semantic segmentation field, showing its capability of outperforming several semantic segmentation models on the task of RGB-D SOD. These facts further confirm that the proposed framework could offer a potential solution for various applications and provide more insight into the cross-modal complementarity task.

📄 PDF Abstract BibTeX arXiv:2008.12134

Code (2)

kerenfu/JLDCF 공식 구현 pytorch
taozh2017/RGBD-SODsurvey

Tasks

object-detectionObject DetectionRGB-D Salient Object DetectionRGB Salient Object DetectionSalient Object DetectionSemantic Segmentation

Similar Papers 제목 키워드 기반

SE2Net: Siamese Edge-Enhancement Network for Salient Object Detection

2019-03-29 · Sanping Zhou, Jimuyang Zhang, Jinjun Wang, Fei Wang 외

Deep convolutional neural network significantly boosted the capability of salient object detection in handling large variations of scenes and object appearances. However, convolution operations seek to generate strong re…

Objectobject-detectionObject DetectionRGB Salient Object Detection+1

Multi-interactive Encoder-decoder Network for RGBT Salient Object Detection

2020-06-05 · Zhengzheng Tu, Zhun Li, Chenglong Li, Yang Lang 외

RGBT salient object detection (SOD) aims to segment the common prominent regions of visible and thermal infrared images. Existing RGBT SOD methods don't fully explore and exploit the potentials of complementarity of diff…

Decoderobject-detectionObject DetectionSalient Object Detection

Dynamic Message Propagation Network for RGB-D Salient Object Detection

2022-06-20 · Baian Chen, Zhilei Chen, Xiaowei Hu, Jun Xu 외

This paper presents a novel deep neural network framework for RGB-D salient object detection by controlling the message passing between the RGB images and depth maps on the feature level and exploring the long-range sema…

object-detectionObject DetectionRGB-D Salient Object DetectionSalient Object Detection

Panoramic Video Salient Object Detection with Ambisonic Audio Guidance

2022-11-26 · Xiang Li, Haoyuan Cao, Shijie Zhao, Junlin Li 외

Video salient object detection (VSOD), as a fundamental computer vision problem, has been extensively discussed in the last decade. However, all existing works focus on addressing the VSOD problem in 2D scenarios. With t…

Objectobject-detectionObject DetectionSalient Object Detection+1

SiaTrans: Siamese Transformer Network for RGB-D Salient Object Detection with Depth Image Classification

2022-07-09 · Xingzhao Jia, Dongye Changlei, Yanjun Peng

RGB-D SOD uses depth information to handle challenging scenes and obtain high-quality saliency maps. Existing state-of-the-art RGB-D saliency detection methods overwhelmingly rely on the strategy of directly fusing depth…

image-classificationImage ClassificationMisinformationobject-detection+5