paper-with-me

홈 › Papers

A Decoding Scheme with Successive Aggregation of Multi-Level Features for Light-Weight Semantic Segmentation

2024-02-17 · Jiwon Yoo, JangWon Lee, Gyeonghwan Kim

Multi-scale architecture, including hierarchical vision transformer, has been commonly applied to high-resolution semantic segmentation to deal with computational complexity with minimum performance loss. In this paper, we propose a novel decoding scheme for semantic segmentation in this regard, which takes multi-level features from the encoder with multi-scale architecture. The decoding scheme based on a multi-level vision transformer aims to achieve not only reduced computational expense but also higher segmentation accuracy, by introducing successive cross-attention in aggregation of the multi-level features. Furthermore, a way to enhance the multi-level features by the aggregated semantics is proposed. The effort is focused on maintaining the contextual consistency from the perspective of attention allocation and brings improved performance with significantly lower computational cost. Set of experiments on popular datasets demonstrates superiority of the proposed scheme to the state-of-the-art semantic segmentation models in terms of computational cost without loss of accuracy, and extensive ablation studies prove the effectiveness of ideas proposed.

📄 PDF Abstract BibTeX arXiv:2402.11201

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Attention 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Multi-Head Attention 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Residual Connection 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…

Similar Papers 제목 키워드 기반

NOMA for Multiple Access Channel and Broadcast Channel in Indoor VLC

2020-08-13 · T. Uday, Abhinav Kumar, L. Natarajan

Orthogonal frequency division multiplexing (OFDM) based non-orthogonal multiple access (NOMA) has increased complexity and reduced spectral efficiency in visible light communications (VLC) NOMA compared to radiofrequency…

BiSPARCs for Unsourced Random Access in Massive MIMO

2023-04-27 · Patrick Agostini, Zoran Utkovski, Slawomir Stanczak

This paper considers the massive MIMO unsourced random access problem in a quasi-static Rayleigh fading setting. The proposed coding scheme is based on a concatenation of a "conventional" channel code (such as, e.g., LDP…

High Throughput Polar Decoding Using Two-Staged Adaptive Successive Cancellation List Decoding

2019-05-22

Polar codes are the first class of capacity-achieving forward error correction (FEC) codes. They have been selected as one of the coding schemes for the 5G communication systems due to their excellent error correction pe…

Structured Superposition of Autoencoders for UEP Codes at Intermediate Blocklengths

2025-08-10 · Vukan Ninkovic, Dejan Vukobratovic arxiv

Unequal error protection (UEP) coding that enables differentiated reliability levels within a transmitted message is essential for modern communication systems. Autoencoder (AE)-based code designs have shown promise in t…

Improved Automorphism Ensemble Decoder for Polar Codes

2024-08-08 · Conference 2024 8 · Seokju Han, Member, IEEE, Bonghoe Kim 외

In this work, we propose an improved automorphism ensemble (AE) decoder for polar codes. With successive cancellation (SC) variant automorphisms, multiple decoding paths in the AE decoder produce their estimates of the…

Decoder