paper-with-me

홈 › Papers

AstroLoc: Robust Space to Ground Image Localizer

2025-02-10 · Gabriele Berton, Alex Stoken, Carlo Masone

Astronauts take thousands of photos of Earth per day from the International Space Station, which, once localized on Earth's surface, are used for a multitude of tasks, ranging from climate change research to disaster management. The localization process, which has been performed manually for decades, has recently been approached through image retrieval solutions: given an astronaut photo, find its most similar match among a large database of geo-tagged satellite images, in a task called Astronaut Photography Localization (APL). Yet, existing APL approaches are trained only using satellite images, without taking advantage of the millions open-source astronaut photos. In this work we present the first APL pipeline capable of leveraging astronaut photos for training. We first produce full localization information for 300,000 manually weakly labeled astronaut photos through an automated pipeline, and then use these images to train a model, called AstroLoc. AstroLoc learns a robust representation of Earth's surface features through two losses: astronaut photos paired with their matching satellite counterparts in a pairwise loss, and a second loss on clusters of satellite imagery weighted by their relevance to astronaut photography via unsupervised mining. We find that AstroLoc achieves a staggering 35% average improvement in recall@1 over previous SOTA, pushing the limits of existing datasets with a recall@100 consistently over 99%. Finally, we note that AstroLoc, without any fine-tuning, provides excellent results for related tasks like the lost-in-space satellite problem and historical space imagery localization.

📄 PDF Abstract BibTeX arXiv:2502.07003

Code (0)

등록된 구현이 없습니다.

Tasks

Image Retrieval

Similar Papers 제목 키워드 기반

Opportunistic Cardiac Health Assessment: Estimating Phenotypes from Localizer MRI through Multi-Modal Representations

2026-03-13 · Busra Nur Zeybek, Özgün Turgut, Yundi Zhang, Jiazhen Pan 외 arxiv

Cardiovascular diseases are the leading cause of death. Cardiac phenotypes (CPs), e.g., ejection fraction, are the gold standard for assessing cardiac health, but they are derived from cine cardiac magnetic resonance ima…

Improving Weakly-Supervised Object Localization Using Adversarial Erasing and Pseudo Label

2024-04-15 · Byeongkeun Kang, Sinhae Cha, Yeejin Lee

Weakly-supervised learning approaches have gained significant attention due to their ability to reduce the effort required for human annotations in training neural networks. This paper investigates a framework for weakly…

ObjectObject LocalizationPseudo LabelWeakly-supervised Learning+1

Leveraging Uncertainty for Deep Interpretable Classification and Weakly-Supervised Segmentation of Histology Images

2022-05-12 · Soufiane Belharbi, Jérôme Rony, Jose Dolz, Ismail Ben Ayed 외

Trained using only image class label, deep weakly supervised methods allow image classification and ROI segmentation for interpretability. Despite their success on natural images, they face several challenges over histol…

image-classificationImage ClassificationSegmentationWeakly supervised segmentation

Talking Points: Describing and Localizing Pixels

2025-10-16 · Matan Rusanovsky, Shimon Malnick, Shai Avidan arxiv

Vision-language models have achieved remarkable success in cross-modal understanding. Yet, these models remain limited to object-level or region-level grounding, lacking the capability for pixel-precise keypoint comprehe…

Self-Chained Image-Language Model for Video Localization and Question Answering

2023-05-11 · NeurIPS 2023 11 · Shoubin Yu, Jaemin Cho, Prateek Yadav, Mohit Bansal

Recent studies have shown promising results on utilizing large pre-trained image-language models for video question answering. While these image-language models can efficiently bootstrap the representation learning of vi…

Language ModelingLanguage ModellingQuestion AnsweringRepresentation Learning+3