paper-with-me

Papers

RRNet: Relational Reasoning Network with Parallel Multi-scale Attention for Salient Object Detection in Optical Remote Sensing Images

2021-10-27 · Runmin Cong, Yumo Zhang, Leyuan Fang, Jun Li, Yao Zhao, Sam Kwong

Salient object detection (SOD) for optical remote sensing images (RSIs) aims at locating and extracting visually distinctive objects/regions from the optical RSIs. Despite some saliency models were proposed to solve the intrinsic problem of optical RSIs (such as complex background and scale-variant objects), the accuracy and completeness are still unsatisfactory. To this end, we propose a relational reasoning network with parallel multi-scale attention for SOD in optical RSIs in this paper. The relational reasoning module that integrates the spatial and the channel dimensions is designed to infer the semantic relationship by utilizing high-level encoder features, thereby promoting the generation of more complete detection results. The parallel multi-scale attention module is proposed to effectively restore the detail information and address the scale variation of salient objects by using the low-level features refined by multi-scale attention. Extensive experiments on two datasets demonstrate that our proposed RRNet outperforms the existing state-of-the-art SOD competitors both qualitatively and quantitatively.

📄 PDF Abstract BibTeX arXiv:2110.14223

Code (2)

2023-MindSpore-1/ms-code-175 mindspore
rmcong/RRNet_TGRS2021 pytorch

Tasks

object-detectionObject DetectionRelational ReasoningSalient Object Detection

Similar Papers 제목 키워드 기반

Explicit Relational Reasoning Network for Scene Text Detection

2024-12-19 · Yuchen Su, Zhineng Chen, Yongkun Du, Zhilong Ji 외

Connected component (CC) is a proper text shape representation that aligns with human reading intuition. However, CC-based text detection methods have recently faced a developmental bottleneck that their time-consuming p…

DecoderRelational ReasoningScene Text DetectionText Detection

Correlational Neural Networks

2015-04-27 · Sarath Chandar, Mitesh M. Khapra, Hugo Larochelle, Balaraman Ravindran

Common Representation Learning (CRL), wherein different descriptions (or views) of the data are embedded in a common subspace, is receiving a lot of attention recently. Two popular paradigms here are Canonical Correlatio…

Representation LearningTransfer Learning

CorrNet+: Sign Language Recognition and Translation via Spatial-Temporal Correlation

2024-04-17 · Lianyu Hu, Wei Feng, Liqing Gao, Zekang Liu 외

In sign language, the conveyance of human body trajectories predominantly relies upon the coordinated movements of hands and facial expressions across successive frames. Despite the recent advancements of sign language u…

Pose EstimationSign Language RecognitionSign Language TranslationTranslation

Continuous Sign Language Recognition with Correlation Network

2023-03-06 · CVPR 2023 1 · Lianyu Hu, Liqing Gao, Zekang Liu, Wei Feng

Human body trajectories are a salient cue to identify actions in the video. Such body trajectories are mainly conveyed by hands and face across consecutive frames in sign language. However, current methods in continuous …

Sign Language Recognition

CorrNet3D: Unsupervised End-to-end Learning of Dense Correspondence for 3D Point Clouds

2020-12-31 · CVPR 2021 1 · Yiming Zeng, Yue Qian, Zhiyu Zhu, Junhui Hou 외

Motivated by the intuition that one can transform two aligned point clouds to each other more easily and meaningfully than a misaligned pair, we propose CorrNet3D -- the first unsupervised and end-to-end deep learning-ba…

3D Dense Shape Correspondence