CVEGAN: A Perceptually-inspired GAN for Compressed Video Enhancement
We propose a new Generative Adversarial Network for Compressed Video quality Enhancement (CVEGAN). The CVEGAN generator benefits from the use of a novel Mul2Res block (with multiple levels of residual learning branches), an enhanced residual non-local block (ERNB) and an enhanced convolutional block attention module (ECBAM). The ERNB has also been employed in the discriminator to improve the representational capability. The training strategy has also been re-designed specifically for video compression applications, to employ a relativistic sphere GAN (ReSphereGAN) training methodology together with new perceptual loss functions. The proposed network has been fully evaluated in the context of two typical video compression enhancement tools: post-processing (PP) and spatial resolution adaptation (SRA). CVEGAN has been fully integrated into the MPEG HEVC video coding test model (HM16.20) and experimental results demonstrate significant coding gains (up to 28% for PP and 38% for SRA compared to the anchor) over existing state-of-the-art architectures for both coding tools across multiple datasets.
Code (1)
Tasks
Generative Adversarial NetworkVideo CompressionVideo EnhancementMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Enhancing VVC with Deep Learning based Multi-Frame Post-Processing
This paper describes a CNN-based multi-frame post-processing approach based on a perceptually-inspired Generative Adversarial Network architecture, CVEGAN. This method has been integrated with the Versatile Video Coding …
Deep LearningGenerative Adversarial NetworkImage CompressionPerceptually-inspired super-resolution of compressed videos
Spatial resolution adaptation is a technique which has often been employed in video compression to enhance coding efficiency. This approach encodes a lower resolution version of the input video and reconstructs the origi…
Generative Adversarial NetworkSuper-ResolutionVideo CompressionLearnt Deep Hyperparameter selection in Adversarial Training for compressed video enhancement with perceptual critic
Image based Deep Feature Quality Metrics (DFQMs) have been shown to better correlate with subjective perceptual scores over traditional metrics. The fundamental focus of these DFQMs is to exploit internal representations…
feature selectionVideo EnhancementValid Information Guidance Network for Compressed Video Quality Enhancement
In recent years deep learning methods have shown great superiority in compressed video quality enhancement tasks. Existing methods generally take the raw video as the ground truth and extract practical information from c…
validA Diffusion Model Based Quality Enhancement Method for HEVC Compressed Video
Video post-processing methods can improve the quality of compressed videos at the decoder side. Most of the existing methods need to train corresponding models for compressed videos with different quantization parameters…
DecoderQuantization