paper-with-me

홈 › Papers

DarSwin-Unet: Distortion Aware Encoder-Decoder Architecture

2024-07-24 · Akshaya Athwale, Ichrak Shili, Émile Bergeron, Ola Ahmad, Jean-François Lalonde

Wide-angle fisheye images are becoming increasingly common for perception tasks in applications such as robotics, security, and mobility (e.g. drones, avionics). However, current models often either ignore the distortions in wide-angle images or are not suitable to perform pixel-level tasks. In this paper, we present an encoder-decoder model based on a radial transformer architecture that adapts to distortions in wide-angle lenses by leveraging the physical characteristics defined by the radial distortion profile. In contrast to the original model, which only performs classification tasks, we introduce a U-Net architecture, DarSwin-Unet, designed for pixel level tasks. Furthermore, we propose a novel strategy that minimizes sparsity when sampling the image for creating its input tokens. Our approach enhances the model capability to handle pixel-level tasks in wide-angle fisheye images, making it more effective for real-world applications. Compared to other baselines, DarSwin-Unet achieves the best results across different datasets, with significant gains when trained on bounded levels of distortions (very low, low, medium, and high) and tested on all, including out-of-distribution distortions. We demonstrate its performance on depth estimation and show through extensive experiments that DarSwin-Unet can perform zero-shot adaptation to unseen distortions of different wide-angle lenses.

📄 PDF Abstract BibTeX arXiv:2407.17328

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDepth Estimation

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
U-Net 설명 없음

Similar Papers 제목 키워드 기반

DarSwin: Distortion Aware Radial Swin Transformer

2023-04-19 · ICCV 2023 1 · Akshaya Athwale, Arman Afrasiyabi, Justin Lagüe, Ichrak Shili 외

Wide-angle lenses are commonly used in perception tasks requiring a large field of view. Unfortunately, these lenses produce significant distortions, making conventional models that ignore the distortion effects unable t…

Depth Estimation

Context Aware 3D UNet for Brain Tumor Segmentation

2020-10-25 · Parvez Ahmad, Saqib Qamar, Linlin Shen, Adnan Saeed

Deep convolutional neural network (CNN) achieves remarkable performance for medical image analysis. UNet is the primary source in the performance of 3D CNN architectures for medical imaging tasks, including brain tumor s…

Brain Tumor SegmentationDecoderMedical Image AnalysisSegmentation+1

FCSN: Global Context Aware Segmentation by Learning the Fourier Coefficients of Objects in Medical Images

2022-07-29 · Young Seok Jeon, Hongfei Yang, Mengling Feng

The encoder-decoder model is a commonly used Deep Neural Network (DNN) model for medical image segmentation. Conventional encoder-decoder models make pixel-wise predictions focusing heavily on local patterns around the p…

DecoderImage SegmentationMedical Image SegmentationSegmentation+1

D-TrAttUnet: Dual-Decoder Transformer-Based Attention Unet Architecture for Binary and Multi-classes Covid-19 Infection Segmentation

2023-03-27 · Fares Bougourzi, Cosimo Distante, Fadi Dornaika, Abdelmalik Taleb-Ahmed

In the last three years, the world has been facing a global crisis caused by Covid-19 pandemic. Medical imaging has been playing a crucial role in the fighting against this disease and saving the human lives. Indeed, CT-…

DecoderSegmentation

GCA-SUNet: A Gated Context-Aware Swin-UNet for Exemplar-Free Counting

2024-09-18 · Yuzhe Wu, Yipeng Xu, Tianyu Xu, Jialu Zhang 외

Exemplar-Free Counting aims to count objects of interest without intensive annotations of objects or exemplars. To achieve this, we propose a Gated Context-Aware Swin-UNet (GCA-SUNet) to directly map an input image to th…

DecoderExemplar-FreeExemplar-Free CountingObject Counting+1