The Synthinel-1 dataset: a collection of high resolution synthetic overhead imagery for building segmentation
Recently deep learning - namely convolutional neural networks (CNNs) - have yielded impressive performance for the task of building segmentation on large overhead (e.g., satellite) imagery benchmarks. However, these benchmark datasets only capture a small fraction of the variability present in real-world overhead imagery, limiting the ability to properly train, or evaluate, models for real-world application. Unfortunately, developing a dataset that captures even a small fraction of real-world variability is typically infeasible due to the cost of imagery, and manual pixel-wise labeling of the imagery. In this work we develop an approach to rapidly and cheaply generate large and diverse virtual environments from which we can capture synthetic overhead imagery for training segmentation CNNs. Using this approach, generate and publicly-release a collection of synthetic overhead imagery - termed Synthinel-1 with full pixel-wise building labels. We use several benchmark dataset to demonstrate that Synthinel-1 is consistently beneficial when used to augment real-world training imagery, especially when CNNs are tested on novel geographic locations or conditions.
Code (1)
Similar Papers 제목 키워드 기반
SYNC: A Copula based Framework for Generating Synthetic Data from Aggregated Sources
A synthetic dataset is a data object that is generated programmatically, and it may be valuable to creating a single dataset from multiple sources when direct collection is difficult or costly. Although it is a fundament…
Feature EngineeringSynthetic Data GenerationA Masked Face Classification Benchmark on Low-Resolution Surveillance Images
We propose a novel image dataset focused on tiny faces wearing face masks for mask classification purposes, dubbed Small Face MASK (SF-MASK), composed of a collection made from 20k low-resolution images exported from div…
ClassificationMulti-class ClassificationSynRS3D: A Synthetic Dataset for Global 3D Semantic Understanding from Monocular Remote Sensing Imagery
Global semantic 3D understanding from single-view high-resolution remote sensing (RS) imagery is crucial for Earth Observation (EO). However, this task faces significant challenges due to the high costs of annotations an…
Domain AdaptationEarth ObservationSynthetic Data GenerationUnsupervised Domain AdaptationFLUXSynID: A Framework for Identity-Controlled Synthetic Face Generation with Document and Live Images
Synthetic face datasets are increasingly used to overcome the limitations of real-world biometric data, including privacy concerns, demographic imbalance, and high collection costs. However, many existing methods lack fi…
DiversityFace GenerationFace RecognitionDelving into High-Quality Synthetic Face Occlusion Segmentation Datasets
This paper performs comprehensive analysis on datasets for occlusion-aware face segmentation, a task that is crucial for many downstream applications. The collection and annotation of such datasets are time-consuming and…
SegmentationSynthetic Data GenerationVocal Bursts Intensity Prediction