Multiple instance dense connected convolution neural network for aerial image scene classification
With the development of deep learning, many state-of-the-art natural image scene classification methods have demonstrated impressive performance. While the current convolution neural network tends to extract global features and global semantic information in a scene, the geo-spatial objects can be located at anywhere in an aerial image scene and their spatial arrangement tends to be more complicated. One possible solution is to preserve more local semantic information and enhance feature propagation. In this paper, an end to end multiple instance dense connected convolution neural network (MIDCCNN) is proposed for aerial image scene classification. First, a 23 layer dense connected convolution neural network (DCCNN) is built and served as a backbone to extract convolution features. It is capable of preserving middle and low level convolution features. Then, an attention based multiple instance pooling is proposed to highlight the local semantics in an aerial image scene. Finally, we minimize the loss between the bag-level predictions and the ground truth labels so that the whole framework can be trained directly. Experiments on three aerial image datasets demonstrate that our proposed methods can outperform current baselines by a large margin.
Code (0)
등록된 구현이 없습니다.
Tasks
General ClassificationScene ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
A multiple-instance densely-connected ConvNet for aerial scene classification
In contrast with nature scenes, aerial scenes are often composed of many objects crowdedly distributed on the surface in bird’s view, the description of which usually demands more discriminative features as well as local…
Aerial Scene ClassificationClassificationImage ClassificationMultiple Instance Learning+2Dense-TNT: Efficient Vehicle Type Classification Neural Network Using Satellite Imagery
Accurate vehicle type classification serves a significant role in the intelligent transportation system. It is critical for ruler to understand the road conditions and usually contributive for the traffic light control s…
ClassificationVocal Bursts Type PredictionSalient Instance Segmentation via Subitizing and Clustering
The goal of salient region detection is to identify the regions of an image that attract the most attention. Many methods have achieved state-of-the-art performance levels on this task. Recently, salient instance segment…
ClusteringInstance SegmentationSegmentationSemantic SegmentationLoD-Loc v3: Generalized Aerial Localization in Dense Cities using Instance Silhouette Alignment
We present LoD-Loc v3, a novel method for generalized aerial visual localization in dense urban environments. While prior work LoD-Loc v2 achieves localization through semantic building silhouette alignment with low-deta…
Synthetic Data GenerationZero-shot GeneralizationInstance SegmentationVisual LocalizationAn Accurate Car Counting in Aerial Images Based on Convolutional Neural Networks
This paper proposes a simple and effective single-shot detector model to detect and count cars in aerial images. The proposed model, called heatmap learner convolutional neural network (HLCNN), is used to predict the h…
Data AugmentationObject Counting