paper-with-me

홈 › Papers

UAVid: A Semantic Segmentation Dataset for UAV Imagery

2018-10-24 · Ye Lyu, George Vosselman, Gui-Song Xia, Alper Yilmaz, Michael Ying Yang

Semantic segmentation has been one of the leading research interests in computer vision recently. It serves as a perception foundation for many fields, such as robotics and autonomous driving. The fast development of semantic segmentation attributes enormously to the large scale datasets, especially for the deep learning related methods. There already exist several semantic segmentation datasets for comparison among semantic segmentation methods in complex urban scenes, such as the Cityscapes and CamVid datasets, where the side views of the objects are captured with a camera mounted on the driving car. There also exist semantic labeling datasets for the airborne images and the satellite images, where the top views of the objects are captured. However, only a few datasets capture urban scenes from an oblique Unmanned Aerial Vehicle (UAV) perspective, where both of the top view and the side view of the objects can be observed, providing more information for object recognition. In this paper, we introduce our UAVid dataset, a new high-resolution UAV semantic segmentation dataset as a complement, which brings new challenges, including large scale variation, moving object recognition and temporal consistency preservation. Our UAV dataset consists of 30 video sequences capturing 4K high-resolution images in slanted views. In total, 300 images have been densely labeled with 8 classes for the semantic labeling task. We have provided several deep learning baseline methods with pre-training, among which the proposed Multi-Scale-Dilation net performs the best via multi-scale feature extraction. Our UAVid website and the labeling tool have been published https://uavid.nl/.

📄 PDF Abstract BibTeX arXiv:1810.10438

Code (3)

YeLyuUT/MSDNet 공식 구현 tf
YeLyuUT/UAVidToolKit
enot-autodl/lpcv-2023 pytorch

Tasks

4kAutonomous DrivingObject RecognitionScene UnderstandingSegmentationSemantic SegmentationVideo Semantic Segmentation

Methods 이 논문이 사용한 방법론

CRF Conditional Random Fields or CRFs are a type of probabilistic graph model that take neighboring sample context into account for tasks like classification. Prediction is…

Similar Papers 제목 키워드 기반

U-Net Ensemble for Enhanced Semantic Segmentation in Remote Sensing Imagery

2024-06-08 · Remote Sensing 2024 6 · Ivica Dimitrovski, Vlatko Spasev, Suzana Loshkovska, Ivan Kitanovski

Semantic segmentation of remote sensing imagery stands as a fundamental task within the domains of both remote sensing and computer vision. Its objective is to generate a comprehensive pixel-wise segmentation map of an i…

SegmentationSegmentation Of Remote Sensing ImagerySemantic Segmentation

Bidirectional Multi-scale Attention Networks for Semantic Segmentation of Oblique UAV Imagery

2021-02-05 · Ye Lyu, George Vosselman, Gui-Song Xia, Michael Ying Yang

Semantic segmentation for aerial platforms has been one of the fundamental scene understanding task for the earth observation. Most of the semantic segmentation research focused on scenes captured in nadir view, in which…

Earth ObservationScene UnderstandingSegmentationSemantic Segmentation

Prototype-Based Low Altitude UAV Semantic Segmentation

2026-04-02 · Da Zhang, Gao Junyu, Zhao Zhiyuan arxiv

Semantic segmentation of low-altitude UAV imagery presents unique challenges due to extreme scale variations, complex object boundaries, and limited computational resources on edge devices. Existing transformer-based seg…

Computational EfficiencySemantic Segmentation

Real-time Semantic Segmentation with Context Aggregation Network

2020-11-02 · Michael Ying Yang, Saumya Kumaar, Ye Lyu, Francesco Nex

With the increasing demand of autonomous systems, pixelwise semantic segmentation for visual scene understanding needs to be not only accurate but also efficient for potential real-time applications. In this paper, we pr…

Real-Time Semantic SegmentationScene UnderstandingSegmentationSemantic Segmentation

Zero-Parameter Geometric Gating for Temporally Stable Low-Altitude UAV Video Semantic Segmentation

2026-06-08 · Jingpu Yang, Fengxian Ji, Zhengzhao Lai, Juanfan Wu 외 arxiv

Video semantic segmentation for low-altitude UAVs requires temporal consistency, yet dense optical flow introduces spatially structured noise in the planar regions that dominate aerial imagery. We propose a zero-paramete…

Video Semantic SegmentationSemantic Similarity