paper-with-me

Papers

Discoverability in Satellite Imagery: A Good Sentence is Worth a Thousand Pictures

2020-01-03 · David Noever, Wes Regian, Matt Ciolino, Josh Kalin, Dom Hambrick, Kaye Blankenship

Small satellite constellations provide daily global coverage of the earth's landmass, but image enrichment relies on automating key tasks like change detection or feature searches. For example, to extract text annotations from raw pixels requires two dependent machine learning models, one to analyze the overhead image and the other to generate a descriptive caption. We evaluate seven models on the previously largest benchmark for satellite image captions. We extend the labeled image samples five-fold, then augment, correct and prune the vocabulary to approach a rough min-max (minimum word, maximum description). This outcome compares favorably to previous work with large pre-trained image models but offers a hundred-fold reduction in model size without sacrificing overall accuracy (when measured with log entropy loss). These smaller models provide new deployment opportunities, particularly when pushed to edge processors, on-board satellites, or distributed ground stations. To quantify a caption's descriptiveness, we introduce a novel multi-class confusion or error matrix to score both human-labeled test data and never-labeled images that include bounding box detection but lack full sentence captions. This work suggests future captioning strategies, particularly ones that can enrich the class coverage beyond land use applications and that lessen color-centered and adjacency adjectives ("green", "near", "between", etc.). Many modern language transformers present novel and exploitable models with world knowledge gleaned from training from their vast online corpus. One interesting, but easy example might learn the word association between wind and waves, thus enriching a beach scene with more than just color descriptions that otherwise might be accessed from raw pixels without text annotation.

📄 PDF Abstract BibTeX arXiv:2001.05839

Code (0)

등록된 구현이 없습니다.

Tasks

Change DetectionDescriptiveImage CaptioningSentencetext annotationWorld Knowledge

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Using Satellite Imagery for Good: Detecting Communities in Desert and Mapping Vaccination Activities

2017-05-12 · Anza Shakeel, Mohsen Ali

Deep convolutional neural networks (CNNs) have outperformed existing object recognition and detection algorithms. On the other hand satellite imagery captures scenes that are diverse. This paper describes a deep learning…

Object Recognition

Attention-Based Scattering Network for Satellite Imagery

2022-10-21 · Jason Stock, Chuck Anderson

Multi-channel satellite imagery, from stacked spectral bands or spatiotemporal data, have meaningful representations for various atmospheric properties. Combining these features in an effective manner to create a perform…

Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi

2026-04-11 · Patrick Kage, Pavlos Andreadis arxiv

Vision foundation models have been shown to be effective at processing satellite imagery into representations fit for downstream tasks, however, creating models which operate over multiple spatial resolutions and modes i…

Representation Learning

Cross-View Splatter: Feed-Forward View Synthesis with Georeferenced Images

2026-05-19 · Matias Turkulainen, Akshay Krishnan, Filippo Aleotti, Mohamed Sayed 외 arxiv

We present Cross-View Splatter, a feed-forward method that predicts pixel-aligned Gaussian splats for outdoor scenes captured at ground level AND by satellite. Faithful reconstructions require good camera coverage, but g…

Deepfake Geography: Detecting AI-Generated Satellite Images

2025-11-21 · Mansur Yerzhanuly arxiv

The rapid advancement of generative models such as StyleGAN2 and Stable Diffusion poses a growing threat to the authenticity of satellite imagery, which is increasingly vital for reliable analysis and decision-making acr…

DeepFake Detection