paper-with-me

홈 › Papers

Stage-Aware Feature Alignment Network for Real-Time Semantic Segmentation of Street Scenes

2022-03-08 · Xi Weng, Yan Yan, Si Chen, Jing-Hao Xue, Hanzi Wang

Over the past few years, deep convolutional neural network-based methods have made great progress in semantic segmentation of street scenes. Some recent methods align feature maps to alleviate the semantic gap between them and achieve high segmentation accuracy. However, they usually adopt the feature alignment modules with the same network configuration in the decoder and thus ignore the different roles of stages of the decoder during feature aggregation, leading to a complex decoder structure. Such a manner greatly affects the inference speed. In this paper, we present a novel Stage-aware Feature Alignment Network (SFANet) based on the encoder-decoder structure for real-time semantic segmentation of street scenes. Specifically, a Stage-aware Feature Alignment module (SFA) is proposed to align and aggregate two adjacent levels of feature maps effectively. In the SFA, by taking into account the unique role of each stage in the decoder, a novel stage-aware Feature Enhancement Block (FEB) is designed to enhance spatial details and contextual information of feature maps from the encoder. In this way, we are able to address the misalignment problem with a very simple and efficient multi-branch decoder structure. Moreover, an auxiliary training strategy is developed to explicitly alleviate the multi-scale object problem without bringing additional computational costs during the inference phase. Experimental results show that the proposed SFANet exhibits a good balance between accuracy and speed for real-time semantic segmentation of street scenes. In particular, based on ResNet-18, SFANet respectively obtains 78.1% and 74.7% mean of class-wise Intersection-over-Union (mIoU) at inference speeds of 37 FPS and 96 FPS on the challenging Cityscapes and CamVid test datasets by using only a single GTX 1080Ti GPU.

📄 PDF Abstract BibTeX arXiv:2203.04031

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderGPUReal-Time Semantic SegmentationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Scale-aware Two-stage High Dynamic Range Imaging

2023-03-12 · Hui Li, Xuyang Yao, Wuyuan Xie, Miaohui Wang

Deep high dynamic range (HDR) imaging as an image translation issue has achieved great performance without explicit optical flow alignment. However, challenges remain over content association ambiguities especially cause…

Optical Flow EstimationVocal Bursts Intensity PredictionVocal Bursts Valence Prediction

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding

2026-05-22 · Jiahe Meng, Weiming Zeng, Yueyang Li, Bo Chai 외 arxiv

Electroencephalography (EEG) visual decoding remains challenging due to the modality gap between low-SNR neural signals and highly structured vision--language spaces, making direct cross-modal alignment unstable. To addr…

UPDA: Unsupervised Progressive Domain Adaptation for No-Reference Point Cloud Quality Assessment

2026-02-12 · Bingxu Xie, Fang Zhou, Jincan Wu, Yonghui Liu 외 arxiv

While no-reference point cloud quality assessment (NR-PCQA) approaches have achieved significant progress over the past decade, their performance often degrades substantially when a distribution gap exists between the tr…

Point Cloud Quality AssessmentDomain Adaptation

Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation

2024-08-14 · Yubin Cho, Hyunwoo Yu, Suk-Ju Kang

Referring segmentation aims to segment a target object related to a natural language expression. Key challenges of this task are understanding the meaning of complex and ambiguous language expressions and determining the…

cross-modal alignmentImage SegmentationSemantic Segmentation

HairFIT: Pose-Invariant Hairstyle Transfer via Flow-based Hair Alignment and Semantic-Region-Aware Inpainting

2022-06-17 · Chaeyeon Chung, Taewoo Kim, Hyelin Nam, Seunghwan Choi 외

Hairstyle transfer is the task of modifying a source hairstyle to a target one. Although recent hairstyle transfer models can reflect the delicate features of hairstyles, they still have two major limitations. First, the…

Optical Flow Estimation