paper-with-me

Papers

Cross-domain heterogeneous residual network for single image super-resolution

2022-05-26 · Neural Networks 2022 5 · Li Ji, Qinghui Zhu, Yongqin Zhang, Juanjuan Yin, Ruyi Wei, Jinsheng Xiao, Deqiang Xiao, Guoying Zhao

Single image super-resolution is an ill-posed problem, whose purpose is to acquire a high-resolution image from its degraded observation. Existing deep learning-based methods are compromised on their performance and speed due to the heavy design (i.e., huge model size) of networks. In this paper, we propose a novel high-performance cross-domain heterogeneous residual network for super-resolved image reconstruction. Our network models heterogeneous residuals between different feature layers by hierarchical residual learning. In outer residual learning, dual-domain enhancement modules extract the frequency-domain information to reinforce the space-domain features of network mapping. In middle residual learning, wide-activated residual-in-residual dense blocks are constructed by concatenating the outputs from previous blocks as the inputs into all subsequent blocks for better parameter efficacy. In inner residual learning, wide-activated residual attention blocks are introduced to capture direction- and location-aware feature maps. The proposed method was evaluated on four benchmark datasets, indicating that it can construct the high-quality super-resolved images and achieve the state-of-the-art performance.

📄 PDF Abstract BibTeX

Code (1)

zhangyongqin/HRN pytorch

Tasks

Image ReconstructionImage Super-ResolutionSuper-Resolution

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Residual Cross-Modal Fusion Networks for Audio-Visual Navigation

2026-01-11 · Yi Wang, Yinfeng Yu, Bin Ren arxiv

Audio-visual embodied navigation aims to enable an agent to autonomously localize and reach a sound source in unseen 3D environments by leveraging auditory cues. The key challenge of this task lies in effectively modelin…

Domain GeneralizationVisual Navigation

Embedded Heterogeneous Attention Transformer for Cross-lingual Image Captioning

2023-07-19 · Zijie Song, Zhenzhen Hu, Yuanen Zhou, Ye Zhao 외

Cross-lingual image captioning is a challenging task that requires addressing both cross-lingual and cross-modal obstacles in multimedia analysis. The crucial issue in this task is to model the global and the local match…

Image Captioning

Residual-Space Evolutionary Optimization via Flow-based Generative Models

2026-06-18 · Zhuo Cao, Lena Krieger, Fernanda Nader, Xuan Zhao 외 arxiv

Data editing with generative methods typically requires differentiable objectives and gradient-based search. However, these assumptions break down in flow-based settings, where edits are performed through forward and bac…

Rethinking Cross-Dose PET Denoising: Mitigating Averaging Effects via Residual Noise Learning

2026-04-18 · Yichao Liu, Zongru Shao, Yueyang Teng, Junwen Guo arxiv

Cross-dose denoising for low-dose positron emission tomography (LDPET) has been proposed to address the limited generalization of models trained at a single noise level. However, neural networks trained on a specific dos…

Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

2026-05-31 · Ziyue Lin, Jiahe Hou, Hongyu Xia, Xinrui Xie 외 arxiv

We propose Decoupled Residual Denoising Diffusion models (DRDD) for unified and data-efficient image-to-image (I2I) translation. While diffusion models have advanced I2I translation in terms of quality and diversity, we …

Image-to-Image Translation