paper-with-me

Papers

RDFNet: RGB-D Multi-Level Residual Feature Fusion for Indoor Semantic Segmentation

2017-10-01 · ICCV 2017 10 · Seong-Jin Park, Ki-Sang Hong, Seungyong Lee

In multi-class indoor semantic segmentation using RGB-D data, it has been shown that incorporating depth feature into RGB feature is helpful to improve segmentation accuracy. However, previous studies have not fully exploited the potentials of multi-modal feature fusion, e.g., simply concatenating RGB and depth features or averaging RGB and depth score maps. To learn the optimal fusion of multi-modal features, this paper presents a novel network that extends the core idea of residual learning to RGB-D semantic segmentation. Our network effectively captures multi-level RGB-D CNN features by including multi-modal feature fusion blocks and multi-level feature refinement blocks. Feature fusion blocks learn residual RGB and depth features and their combinations to fully exploit the complementary characteristics of RGB and depth data. Feature refinement blocks learn the combination of fused features from multiple levels to enable high-resolution prediction. Our network can efficiently train discriminative multi-level features from each modality end-to-end by taking full advantage of skip-connections. Our comprehensive experiments demonstrate that the proposed architecture achieves the state-of-the-art accuracy on two challenging RGB-D indoor datasets, NYUDv2 and SUN RGB-D.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

Deep feature fusion for self-supervised monocular depth prediction

2020-05-16 · Vinay Kaushik, Brejesh lall

Recent advances in end-to-end unsupervised learning has significantly improved the performance of monocular depth prediction and alleviated the requirement of ground truth depth. Although a plethora of work has been done…

DecoderDepth EstimationDepth Prediction

RDFNet: Regional Dynamic FISTA-Net for Spectral Snapshot Compressive Imaging

2023-02-06 · Shiyun Zhou, Tingfa Xu, Shaocong Dong, Jianan Li

Deep convolutional neural networks have recently shown promising results in compressive spectral reconstruction. Previous methods, however, usually adopt a single mapping function for sparse representation. Considering t…

Spectral Reconstruction

TMFNet: Two-Stream Multi-Channels Fusion Networks for Color Image Operation Chain Detection

2024-09-12 · Yakun Niu, Lei Tan, Lei Zhang, Xianyu Zuo

Image operation chain detection techniques have gained increasing attention recently in the field of multimedia forensics. However, existing detection methods suffer from the generalization problem. Moreover, the channel…

Image Operation Chain Detection

Residual Spatial Fusion Network for RGB-Thermal Semantic Segmentation

2023-06-17 · Ping Li, Junjie Chen, Binbin Lin, Xianghua Xu

Semantic segmentation plays an important role in widespread applications such as autonomous driving and robotic sensing. Traditional methods mostly use RGB images which are heavily affected by lighting conditions, \eg, d…

Autonomous DrivingSaliency DetectionSegmentationSemantic Segmentation+1

Iterative Residual Cross-Attention Mechanism: An Integrated Approach for Audio-Visual Navigation Tasks

2025-09-30 · Hailong Zhang, Yinfeng Yu, Liejun Wang, Fuchun Sun 외 arxiv

Audio-visual navigation represents a significant area of research in which intelligent agents utilize egocentric visual and auditory perceptions to identify audio targets. Conventional navigation methodologies typically …

Reinforcement LearningVisual Navigation