paper-with-me

홈 › Papers

OmniSat: Self-Supervised Modality Fusion for Earth Observation

2024-04-12 · Guillaume Astruc, Nicolas Gonthier, Clement Mallet, Loic Landrieu

The diversity and complementarity of sensors available for Earth Observations (EO) calls for developing bespoke self-supervised multimodal learning approaches. However, current multimodal EO datasets and models typically focus on a single data type, either mono-date images or time series, which limits their impact. To address this issue, we introduce OmniSat, a novel architecture able to merge diverse EO modalities into expressive features without labels by exploiting their alignment. To demonstrate the advantages of our approach, we create two new multimodal datasets by augmenting existing ones with new modalities. As demonstrated for three downstream tasks -- forestry, land cover classification, and crop mapping -- OmniSat can learn rich representations without supervision, leading to state-of-the-art performances in semi- and fully supervised settings. Furthermore, our multimodal pretraining scheme improves performance even when only one modality is available for inference. The code and dataset are available at https://github.com/gastruc/OmniSat.

📄 PDF Abstract BibTeX arXiv:2404.08351

Code (1)

gastruc/omnisat 공식 구현 pytorch

Tasks

DiversityEarth ObservationLand Cover ClassificationSelf-Supervised Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

OmniSAT: Compact Action Token, Faster Auto Regression

2025-10-08 · Huaihai Lyu, Chaofan Chen, Senwei Xie, Pengwei Wang 외 arxiv

Existing Vision-Language-Action (VLA) models can be broadly categorized into diffusion-based and auto-regressive (AR) approaches: diffusion models capture continuous action distributions but rely on computationally heavy…

Self-supervised Vision Transformers for Joint SAR-optical Representation Learning

2022-04-11 · Yi Wang, Conrad M Albrecht, Xiao Xiang Zhu

Self-supervised learning (SSL) has attracted much interest in remote sensing and earth observation due to its ability to learn task-agnostic representations without human annotation. While most of the existing SSL works …

Data AugmentationEarth ObservationRepresentation LearningSelf-Supervised Learning

MAESTRO: Masked AutoEncoders for Multimodal, Multitemporal, and Multispectral Earth Observation Data

2025-08-14 · Antoine Labatie, Michael Vaccaro, Nina Lardiere, Anatol Garioud 외 arxiv

Self-supervised learning holds great promise for remote sensing, but standard self-supervised methods must be adapted to the unique characteristics of Earth observation data. We take a step in this direction by conductin…

Self-Supervised Learning

Unsupervised Hyperspectral and Multispectral Image Fusion via Self-Supervised Modality Decoupling

2024-12-06 · Songcheng Du, Yang Zou, Zixu Wang, Xingyuan Li 외

Hyperspectral and Multispectral Image Fusion (HMIF) aims to fuse low-resolution hyperspectral images (LR-HSIs) and high-resolution multispectral images (HR-MSIs) to reconstruct high spatial and high spectral resolution i…

AllSelf-Supervised LearningSuper-Resolutionvalid

Self-MI: Efficient Multimodal Fusion via Self-Supervised Multi-Task Learning with Auxiliary Mutual Information Maximization

2023-11-07 · Cam-Van Thi Nguyen, Ngoc-Hoa Thi Nguyen, Duc-Trong Le, Quang-Thuy Ha

Multimodal representation learning poses significant challenges in capturing informative and distinct features from multiple modalities. Existing methods often struggle to exploit the unique characteristics of each modal…

Multi-Task LearningRepresentation LearningSelf-Supervised Learning