paper-with-me

Papers

ESC-MISR: Enhancing Spatial Correlations for Multi-Image Super-Resolution in Remote Sensing

2024-11-07 · Zhihui Zhang, Jinhui Pang, Jianan Li, Xiaoshuai Hao

Multi-Image Super-Resolution (MISR) is a crucial yet challenging research task in the remote sensing community. In this paper, we address the challenging task of Multi-Image Super-Resolution in Remote Sensing (MISR-RS), aiming to generate a High-Resolution (HR) image from multiple Low-Resolution (LR) images obtained by satellites. Recently, the weak temporal correlations among LR images have attracted increasing attention in the MISR-RS task. However, existing MISR methods treat the LR images as sequences with strong temporal correlations, overlooking spatial correlations and imposing temporal dependencies. To address this problem, we propose a novel end-to-end framework named Enhancing Spatial Correlations in MISR (ESC-MISR), which fully exploits the spatial-temporal relations of multiple images for HR image reconstruction. Specifically, we first introduce a novel fusion module named Multi-Image Spatial Transformer (MIST), which emphasizes parts with clearer global spatial features and enhances the spatial correlations between LR images. Besides, we perform a random shuffle strategy for the sequential inputs of LR images to attenuate temporal dependencies and capture weak temporal correlations in the training stage. Compared with the state-of-the-art methods, our ESC-MISR achieves 0.70dB and 0.76dB cPSNR improvements on the two bands of the PROBA-V dataset respectively, demonstrating the superiority of our method.

📄 PDF Abstract BibTeX arXiv:2411.04706

Code (0)

등록된 구현이 없습니다.

Tasks

Image ReconstructionImage Super-ResolutionSuper-Resolution

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Adam 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Multi-Head Attention 설명 없음
Residual Connection 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…

Similar Papers 제목 키워드 기반

TR-MISR: Multiimage Super-Resolution Based on Feature Fusion With Transformers

2022-02-05 · IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing 2022 2 · Tai An, Xin Zhang, Chunlei Huo, Bin Xue 외

Multiimage super-resolution (MISR), as one of the most promising directions in remote sensing, has become a needy technique in the satellite market. A sequence of images collected by satellites often has plenty of views …

DecoderMulti-Frame Super-ResolutionSuper-Resolution

FL-MISR: Fast Large-Scale Multi-Image Super-Resolution for Computed Tomography Based on Multi-GPU Acceleration

2021-08-09 · Kaicong Sun, Trung-Hieu Tran, Jajnabalkya Guhathakurta, Sven Simon

Multi-image super-resolution (MISR) usually outperforms single-image super-resolution (SISR) under a proper inter-image alignment by explicitly exploiting the inter-image correlation. However, the large computational dem…

Computed Tomography (CT)CPUDistributed OptimizationGPU+2

Deep 3D World Models for Multi-Image Super-Resolution Beyond Optical Flow

2024-01-30 · Luca Savant Aira, Diego Valsesia, Andrea Bordone Molini, Giulia Fracastoro 외

Multi-image super-resolution (MISR) allows to increase the spatial resolution of a low-resolution (LR) acquisition by combining multiple images carrying complementary information in the form of sub-pixel offsets in the s…

Image RegistrationImage Super-ResolutionOptical Flow EstimationSuper-Resolution

Enhanced Self-Supervised Multi-Image Super-Resolution for Camera Array Images

2026-04-08 · Yating Chen, Feng Huang, Xianyu Wu, Jing Wu 외 arxiv

Conventional multi-image super-resolution (MISR) methods, such as burst and video SR, rely on sequential frames from a single camera. Consequently, they suffer from complex image degradation and severe occlusion, increas…

Self-Supervised LearningImage Super-ResolutionImage Restoration

JSTR: Judgment Improves Scene Text Recognition

2024-04-09 · Masato Fujitake

In this paper, we present a method for enhancing the accuracy of scene text recognition tasks by judging whether the image and text match each other. While previous studies focused on generating the recognition results f…

Scene Text Recognition