paper-with-me

Papers

Learning GAN-based Foveated Reconstruction to Recover Perceptually Important Image Features

2021-08-07 · Luca Surace, Marek Wernikowski, Cara Tursun, Karol Myszkowski, Radosław Mantiuk, Piotr Didyk

A foveated image can be entirely reconstructed from a sparse set of samples distributed according to the retinal sensitivity of the human visual system, which rapidly decreases with increasing eccentricity. The use of Generative Adversarial Networks has recently been shown to be a promising solution for such a task, as they can successfully hallucinate missing image information. As in the case of other supervised learning approaches, the definition of the loss function and the training strategy heavily influence the quality of the output. In this work,we consider the problem of efficiently guiding the training of foveated reconstruction techniques such that they are more aware of the capabilities and limitations of the human visual system, and thus can reconstruct visually important image features. Our primary goal is to make the training procedure less sensitive to distortions that humans cannot detect and focus on penalizing perceptually important artifacts. Given the nature of GAN-based solutions, we focus on the sensitivity of human vision to hallucination in case of input samples with different densities. We propose psychophysical experiments, a dataset, and a procedure for training foveated image reconstruction. The proposed strategy renders the generator network flexible by penalizing only perceptually important deviations in the output. As a result, the method emphasized the recovery of perceptually important image features. We evaluated our strategy and compared it with alternative solutions by using a newly trained objective metric, a recent foveated video quality metric, and user experiments. Our evaluations revealed significant improvements in the perceived image reconstruction quality compared with the standard GAN-based training approach.

📄 PDF Abstract BibTeX arXiv:2108.03499

Code (0)

등록된 구현이 없습니다.

Tasks

HallucinationImage ReconstructionSensitivity

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

FoVolNet: Fast Volume Rendering using Foveated Deep Neural Networks

2022-09-20 · David Bauer, Qi Wu, Kwan-Liu Ma

Volume data is found in many important scientific and engineering applications. Rendering this data for visualization at high quality and interactive rates for demanding applications such as virtual reality is still not …

Data VisualizationImage ReconstructionQuantization

Learned Single-Pass Multitasking Perceptual Graphics for Immersive Displays

2024-07-31 · Doğa Yılmaz, Towaki Takikawa, Duygu Ceylan, Kaan Akşit

Immersive displays are advancing rapidly in terms of delivering perceptually realistic images by utilizing emerging perceptual graphics methods such as foveated rendering. In practice, multiple such methods need to be pe…

DenoisingImage Denoising

Training on Foveated Images Improves Robustness to Adversarial Attacks

2023-08-01 · NeurIPS 2023 11

Deep neural networks (DNNs) have been shown to be vulnerable to adversarial attacks -- subtle, perceptually indistinguishable perturbations of inputs that change the response of the model. In the context of vision, we hy…

Towards Metamerism via Foveated Style Transfer

2017-05-29 · ICLR 2019 5 · Arturo Deza, Aditya Jonnalagadda, Miguel Eckstein

The problem of $\textit{visual metamerism}$ is defined as finding a family of perceptually indistinguishable, yet physically different images. In this paper, we propose our NeuroFovea metamer model, a foveated generative…

DecoderMetamerismStyle TransferTexture Synthesis

FOVQA: Blind Foveated Video Quality Assessment

2021-06-24 · Yize Jin, Anjul Patney, Richard Webb, Alan Bovik

Previous blind or No Reference (NR) video quality assessment (VQA) models largely rely on features drawn from natural scene statistics (NSS), but under the assumption that the image statistics are stationary in the spati…

Video CompressionVideo Quality AssessmentVisual Question Answering (VQA)