Super-resolved multi-temporal segmentation with deep permutation-invariant networks
Multi-image super-resolution from multi-temporal satellite acquisitions of a scene has recently enjoyed great success thanks to new deep learning models. In this paper, we go beyond classic image reconstruction at a higher resolution by studying a super-resolved inference problem, namely semantic segmentation at a spatial resolution higher than the one of sensing platform. We expand upon recently proposed models exploiting temporal permutation invariance with a multi-resolution fusion module able to infer the rich semantic information needed by the segmentation task. The model presented in this paper has recently won the AI4EO challenge on Enhanced Sentinel 2 Agriculture.
Code (0)
등록된 구현이 없습니다.
Tasks
Image ReconstructionImage Super-ResolutionSegmentationSemantic SegmentationSuper-ResolutionSimilar Papers 제목 키워드 기반
Permutation invariance and uncertainty in multitemporal image super-resolution
Recent advances have shown how deep neural networks can be extremely effective at super-resolving remote sensing imagery, starting from a multitemporal collection of low-resolution images. However, existing models have n…
Image Super-ResolutionMulti-Frame Super-ResolutionSuper-ResolutionPermutation-Aware Action Segmentation via Unsupervised Frame-to-Segment Alignment
This paper presents an unsupervised transformer-based framework for temporal activity segmentation which leverages not only frame-level cues but also segment-level cues. This is in contrast with previous methods which of…
Action SegmentationDecoderPredictionSegmentation+1Self-Supervised Super-Resolution for Multi-Exposure Push-Frame Satellites
Modern Earth observation satellites capture multi-exposure bursts of push-frame images that can be super-resolved via computational means. In this work, we propose a super-resolution method for such multi-exposure sequen…
Earth ObservationSuper-ResolutionCheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography
Chest radiograph interpretation requires temporal reasoning over prior and current studies, yet most vision-language models are trained on static image-report pairs and lack explicit supervision for modeling longitudinal…
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
We present Inst4DGS, an instance-decomposed 4D Gaussian Splatting (4DGS) approach with long-horizon per-Gaussian trajectories. While dynamic 4DGS has advanced rapidly, instance-decomposed 4DGS remains underexplored, larg…