paper-with-me

홈 › Papers

iSeeBetter: Spatio-Temporal Video Super Resolution using Recurrent-Generative Back-Projection Networks

2020-06-21 · Springer Journal of Computational Visual Media (CVM), Tsinghua University Press 2020 6 · Aman Chadha, John Britto, M. Mani Roja

Recently, learning-based models have enhanced the performance of single-image super-resolution (SISR). However, applying SISR successively to each video frame leads to a lack of temporal coherency. Convolutional neural networks (CNNs) outperform traditional approaches in terms of image quality metrics such as peak signal to noise ratio (PSNR) and structural similarity (SSIM). However, generative adversarial networks (GANs) offer a competitive advantage by being able to mitigate the issue of a lack of finer texture details, usually seen with CNNs when super-resolving at large upscaling factors. We present iSeeBetter, a novel GAN-based spatio-temporal approach to video super-resolution (VSR) that renders temporally consistent super-resolution videos. iSeeBetter extracts spatial and temporal information from the current and neighboring frames using the concept of recurrent back-projection networks as its generator. Furthermore, to improve the "naturality" of the super-resolved image while eliminating artifacts seen with traditional algorithms, we utilize the discriminator from super-resolution generative adversarial network (SRGAN). Although mean squared error (MSE) as a primary loss-minimization objective improves PSNR/SSIM, these metrics may not capture fine details in the image resulting in misrepresentation of perceptual quality. To address this, we use a four-fold (MSE, perceptual, adversarial, and total-variation (TV)) loss function. Our results demonstrate that iSeeBetter offers superior VSR fidelity and surpasses state-of-the-art performance.

📄 PDF Abstract BibTeX

Code (1)

amanchadha/iSeeBetter pytorch

Tasks

Generative Adversarial NetworkImage Super-ResolutionSSIMSuper-ResolutionVideo Super-Resolution

Methods 이 논문이 사용한 방법론

Average Pooling 설명 없음
Batch Normalization 설명 없음
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Residual Connection 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Global Average Pooling Global Average Pooling is a pooling operation designed to replace fully connected layers in classical CNNs. The idea is to generate one feature map for each corresponding…

Similar Papers 제목 키워드 기반

iSeeBetter: Spatio-temporal video super-resolution using recurrent generative back-projection networks

2020-06-13 · Aman Chadha, John Britto, M. Mani Roja

Recently, learning-based models have enhanced the performance of single-image super-resolution (SISR). However, applying SISR successively to each video frame leads to a lack of temporal coherency. Convolutional neural n…

Generative Adversarial NetworkImage Super-ResolutionSSIMSuper-Resolution+1

Deformable 3D Convolution for Video Super-Resolution

2020-04-06 · Xinyi Ying, Longguang Wang, Yingqian Wang, Weidong Sheng 외

The spatio-temporal information among video sequences is significant for video super-resolution (SR). However, the spatio-temporal information cannot be fully used by existing video SR methods since spatial feature extra…

Motion CompensationSuper-ResolutionVideo Super-Resolution

STRPM: A Spatiotemporal Residual Predictive Model for High-Resolution Video Prediction

2022-03-30 · CVPR 2022 1 · Zheng Chang, Xinfeng Zhang, Shanshe Wang, Siwei Ma 외

Although many video prediction methods have obtained good performance in low-resolution (64$\sim$128) videos, predictive models for high-resolution (512$\sim$4K) videos have not been fully explored yet, which are more me…

4kVideo PredictionVocal Bursts Intensity Prediction

Dual-Stream Fusion Network for Spatiotemporal Video Super-Resolution

2021-01-05 · Winter Conference on Applications of Computer Vision (WACV) 2021 1 · Min-Yuan Tseng, Yen-Chung Chen, Yi-Lun Lee, Wei-Sheng Lai 외

Visual data upsampling has been an important research topic for improving the perceptual quality and benefiting various computer vision applications. In recent years, we have witnessed remarkable progresses brought by th…

Image Super-ResolutionSuper-ResolutionVideo Super-Resolution

3DSRnet: Video Super-resolution using 3D Convolutional Neural Networks

2018-12-21 · Soo Ye Kim, Jeongyeon Lim, Taeyoung Na, Munchurl Kim

In video super-resolution, the spatio-temporal coherence between, and among the frames must be exploited appropriately for accurate prediction of the high resolution frames. Although 2D convolutional neural networks (CNN…

Super-ResolutionVideo Super-Resolution