ViSTRA3: Video Coding with Deep Parameter Adaptation and Post Processing
This paper presents a deep learning-based video compression framework (ViSTRA3). The proposed framework intelligently adapts video format parameters of the input video before encoding, subsequently employing a CNN at the decoder to restore their original format and enhance reconstruction quality. ViSTRA3 has been integrated with the H.266/VVC Test Model VTM 14.0, and evaluated under the Joint Video Exploration Team Common Test Conditions. Bj{\o}negaard Delta (BD) measurement results show that the proposed framework consistently outperforms the original VVC VTM, with average BD-rate savings of 1.8% and 3.7% based on the assessment of PSNR and VMAF.
Code (0)
등록된 구현이 없습니다.
Tasks
DecoderVideo CompressionSimilar Papers 제목 키워드 기반
ViSTRA2: Video Coding using Spatial Resolution and Effective Bit Depth Adaptation
We present a new video compression framework (ViSTRA2) which exploits adaptation of spatial resolution and effective bit depth, down-sampling these parameters at the encoder based on perceptual criteria, and up-sampling …
DecoderVideo CompressionA CNN-based Post-Processor for Perceptually-Optimized Immersive Media Compression
In recent years, resolution adaptation based on deep neural networks has enabled significant performance gains for conventional (2D) video codecs. This paper investigates the effectiveness of spatial resolution resamplin…
A Rate-Quality Model for Learned Video Coding
Learned video coding (LVC) has recently achieved superior coding performance. In this paper, we model the rate-quality (R-Q) relationship for learned video coding by a parametric function. We learn a neural network, term…
modelEfficient Adaptation of Neural Network Filter for Video Compression
We present an efficient finetuning methodology for neural-network filters which are applied as a postprocessing artifact-removal step in video coding pipelines. The fine-tuning is performed at encoder side to adapt the n…
Video CompressionDeep Hierarchical Video Compression
Recently, probabilistic predictive coding that directly models the conditional distribution of latent features across successive frames for temporal redundancy removal has yielded promising results. Existing methods usin…
Video Compression