paper-with-me

홈 › Papers

HIIF: Hierarchical Encoding based Implicit Image Function for Continuous Super-resolution

2024-12-04 · CVPR 2025 1 · YuXuan Jiang, Ho Man Kwan, Tianhao Peng, Ge Gao, Fan Zhang, Xiaoqing Zhu, Joel Sole, David Bull

Recent advances in implicit neural representations (INRs) have shown significant promise in modeling visual signals for various low-vision tasks including image super-resolution (ISR). INR-based ISR methods typically learn continuous representations, providing flexibility for generating high-resolution images at any desired scale from their low-resolution counterparts. However, existing INR-based ISR methods utilize multi-layer perceptrons for parameterization in the network; this does not take account of the hierarchical structure existing in local sampling points and hence constrains the representation capability. In this paper, we propose a new \textbf{H}ierarchical encoding based \textbf{I}mplicit \textbf{I}mage \textbf{F}unction for continuous image super-resolution, \textbf{HIIF}, which leverages a novel hierarchical positional encoding that enhances the local implicit representation, enabling it to capture fine details at multiple scales. Our approach also embeds a multi-head linear attention mechanism within the implicit attention network by taking additional non-local information into account. Our experiments show that, when integrated with different backbone encoders, HIIF outperforms the state-of-the-art continuous image super-resolution methods by up to 0.17dB in PSNR. The source code of HIIF will be made publicly available at \url{www.github.com}.

📄 PDF Abstract BibTeX arXiv:2412.03748

Code (0)

등록된 구현이 없습니다.

Tasks

Image Super-ResolutionSuper-Resolution

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Multi-Head Linear Attention Multi-Head Linear Attention is a type of linear multi-head self-attention module, proposed with the Linformer architecture. The…

Similar Papers 제목 키워드 기반

OctField: Hierarchical Implicit Functions for 3D Modeling

2021-11-01 · NeurIPS 2021 12 · Jia-Heng Tang, Weikai Chen, Jie Yang, Bo wang 외

Recent advances in localized implicit functions have enabled neural implicit representation to be scalable to large scenes. However, the regular subdivision of 3D space employed by these approaches fails to take into acc…

HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction

2024-08-03 · Xiaodong Gu, Weihao Yuan, Heng Li, Zilong Dong 외

Neural implicit surface reconstruction has become a new trend in reconstructing a detailed 3D shape from images. In previous methods, however, the 3D scene is only encoded by the MLPs which do not have an explicit 3D str…

Surface Reconstruction

A neural lens for super-resolution biological imaging

2019-07-17 · journal of physics communications 2019 7 · James A. Grant-Jacob, Benita S. Mackay, James A G Baker, Yunhui Xie 외

Visualizing structures smaller than the eye can see has been a driving force in scientific research since the invention of the optical microscope. Here, we use a network of neural networks to create a neural lens that …

Super-Resolution

UltraSR: Spatial Encoding is a Missing Key for Implicit Image Function-based Arbitrary-Scale Super-Resolution

2021-03-23 · Xingqian Xu, Zhangyang Wang, Humphrey Shi

The recent success of NeRF and other related implicit neural representation methods has opened a new path for continuous image representation, where pixel values no longer need to be looked up from stored discrete 2D arr…

NeRFSuper-Resolution

NeuDA: Neural Deformable Anchor for High-Fidelity Implicit Surface Reconstruction

2023-03-04 · CVPR 2023 1 · Bowen Cai, Jinchi Huang, Rongfei Jia, Chengfei Lv 외

This paper studies implicit surface reconstruction leveraging differentiable ray casting. Previous works such as IDR and NeuS overlook the spatial context in 3D space when predicting and rendering the surface, thereby ma…

Surface ReconstructionVocal Bursts Intensity Prediction