paper-with-me

홈 › Papers

Residual Bi-Fusion Feature Pyramid Network for Accurate Single-shot Object Detection

2019-11-27 · Ping-Yang Chen, Jun-Wei Hsieh, Chien-Yao Wang, Hong-Yuan Mark Liao, Munkhjargal Gochoo

State-of-the-art (SoTA) models have improved the accuracy of object detection with a large margin via a FP (feature pyramid). FP is a top-down aggregation to collect semantically strong features to improve scale invariance in both two-stage and one-stage detectors. However, this top-down pathway cannot preserve accurate object positions due to the shift-effect of pooling. Thus, the advantage of FP to improve detection accuracy will disappear when more layers are used. The original FP lacks a bottom-up pathway to offset the lost information from lower-layer feature maps. It performs well in large-sized object detection but poor in small-sized object detection. A new structure "residual feature pyramid" is proposed in this paper. It is bidirectional to fuse both deep and shallow features towards more effective and robust detection for both small-sized and large-sized objects. Due to the "residual" nature, it can be easily trained and integrated to different backbones (even deeper or lighter) than other bi-directional methods. One important property of this residual FP is: accuracy improvement is still found even if more layers are adopted. Extensive experiments on VOC and MS COCO datasets showed the proposed method achieved the SoTA results for highly-accurate and efficient object detection..

📄 PDF Abstract BibTeX arXiv:1911.12051

Code (0)

등록된 구현이 없습니다.

Tasks

Objectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Parallel Residual Bi-Fusion Feature Pyramid Network for Accurate Single-Shot Object Detection

2020-12-03 · Ping-Yang Chen, Ming-Ching Chang, Jun-Wei Hsieh, Yong-Sheng Chen

This paper proposes the Parallel Residual Bi-Fusion Feature Pyramid Network (PRB-FPN) for fast and accurate single-shot object detection. Feature Pyramid (FP) is widely used in recent visual detection, however the top-do…

Multi-Object Trackingobject-detectionObject DetectionReal-Time Object Detection

Structure-Aware Residual Pyramid Network for Monocular Depth Estimation

2019-07-13 · Xiaotian Chen, Xuejin Chen, Zheng-Jun Zha

Monocular depth estimation is an essential task for scene understanding. The underlying structure of objects and stuff in a complex scene is critical to recovering accurate and visually-pleasing depth maps. Global struct…

DecoderDepth EstimationDepth PredictionMonocular Depth Estimation+1

AugFPN: Improving Multi-scale Feature Learning for Object Detection

2019-12-11 · CVPR 2020 6 · Chaoxu Guo, Bin Fan, Qian Zhang, Shiming Xiang 외

Current state-of-the-art detectors typically exploit feature pyramid to detect objects at different scales. Among them, FPN is one of the representative works that build a feature pyramid by multi-scale features summatio…

Objectobject-detectionObject Detection

Pyramid Frequency Network with Spatial Attention Residual Refinement Module for Monocular Depth Estimation

2022-04-05 · Zhengyang Lu, Ying Chen

Deep-learning-based approaches to depth estimation are rapidly advancing, offering superior performance over existing methods. To estimate the depth in real-world scenarios, depth estimation models require the robustness…

Deep LearningDepth EstimationMonocular Depth Estimation

Residual Pyramid Learning for Single-Shot Semantic Segmentation

2019-03-23 · Xiaoyu Chen, Xiaotian Lou, Lianfa Bai, Jing Han

Pixel-level semantic segmentation is a challenging task with a huge amount of computation, especially if the size of input is large. In the segmentation model, apart from the feature extraction, the extra decoder structu…

DecoderSegmentationSemantic Segmentation