paper-with-me

홈 › Papers

Deformable Convolution Based Road Scene Semantic Segmentation of Fisheye Images in Autonomous Driving

2024-07-23 · Anam Manzoor, Aryan Singh, Ganesh Sistu, Reenu Mohandas, Eoin Grua, Anthony Scanlan, Ciarán Eising

This study investigates the effectiveness of modern Deformable Convolutional Neural Networks (DCNNs) for semantic segmentation tasks, particularly in autonomous driving scenarios with fisheye images. These images, providing a wide field of view, pose unique challenges for extracting spatial and geometric information due to dynamic changes in object attributes. Our experiments focus on segmenting the WoodScape fisheye image dataset into ten distinct classes, assessing the Deformable Networks' ability to capture intricate spatial relationships and improve segmentation accuracy. Additionally, we explore different loss functions to address class imbalance issues and compare the performance of conventional CNN architectures with Deformable Convolution-based CNNs, including Vanilla U-Net and Residual U-Net architectures. The significant improvement in mIoU score resulting from integrating Deformable CNNs demonstrates their effectiveness in handling the geometric distortions present in fisheye imagery, exceeding the performance of traditional CNN architectures. This underscores the significant role of Deformable convolution in enhancing semantic segmentation performance for fisheye imagery.

📄 PDF Abstract BibTeX arXiv:2407.16647

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
U-Net 설명 없음
Focus 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Deformable Convolution Deformable convolutions add 2D offsets to the regular grid sampling locations in the standard convolution. It enables free…

Similar Papers 제목 키워드 기반

Restricted Deformable Convolution based Road Scene Semantic Segmentation Using Surround View Cameras

2018-01-02 · Liuyuan Deng, Ming Yang, Hao Li, Tianyi Li 외

Understanding the surrounding environment of the vehicle is still one of the challenges for autonomous driving. This paper addresses 360-degree road scene semantic segmentation using surround view cameras, which are wide…

Autonomous DrivingMulti-Task LearningSegmentationSemantic Segmentation

AD-SAM: Fine-Tuning the Segment Anything Vision Foundation Model for Autonomous Driving Perception

2025-10-30 · Mario Camarena, Het Patel, Fatemeh Nazari, Evangelos Papalexakis 외 arxiv

This paper presents the Autonomous Driving Segment Anything Model (AD-SAM), a fine-tuned vision foundation model for semantic segmentation in autonomous driving (AD). AD-SAM extends the Segment Anything Model (SAM) with …

Semantic SegmentationDomain GeneralizationAutonomous Driving

Spatiotemporal Deformable Scene Graphs for Complex Activity Detection

2021-04-16 · Salman Khan, Fabio Cuzzolin

Long-term complex activity recognition and localisation can be crucial for decision making in autonomous systems such as smart cars and surgical robots. Here we address the problem via a novel deformable, spatiotemporal …

Action DetectionActivity DetectionActivity RecognitionAutonomous Driving+1

Deformable-Heatmap-Segmentation for Automobile Visual Perception

2024-07-10 · Hongyu Jin

Semantic segmentation of road elements in 2D images is a crucial task in the recognition of some static objects such as lane lines and free space. In this paper, we propose DHSNet,which extracts the objects features with…

Semantic Segmentation

Twin Deformable Point Convolutions for Point Cloud Semantic Segmentation in Remote Sensing Scenes

2024-05-30 · Yong-Qiang Mao, Hanbo Bi, Xuexue Li, Kaiqiang Chen 외

Thanks to the application of deep learning technology in point cloud processing of the remote sensing field, point cloud segmentation has become a research hotspot in recent years, which can be applied to real-world 3D, …

Point Cloud SegmentationSegmentationSemantic Segmentation