ViSTRA2: Video Coding using Spatial Resolution and Effective Bit Depth Adaptation
We present a new video compression framework (ViSTRA2) which exploits adaptation of spatial resolution and effective bit depth, down-sampling these parameters at the encoder based on perceptual criteria, and up-sampling at the decoder using a deep convolution neural network. ViSTRA2 has been integrated with the reference software of both the HEVC (HM 16.20) and VVC (VTM 4.01), and evaluated under the Joint Video Exploration Team Common Test Conditions using the Random Access configuration. Our results show consistent and significant compression gains against HM and VVC based on Bj{\o}negaard Delta measurements, with average BD-rate savings of 12.6% (PSNR) and 19.5% (VMAF) over HM and 5.5% (PSNR) and 8.6% (VMAF) over VTM.
Code (0)
등록된 구현이 없습니다.
Tasks
DecoderVideo CompressionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
ViSTRA3: Video Coding with Deep Parameter Adaptation and Post Processing
This paper presents a deep learning-based video compression framework (ViSTRA3). The proposed framework intelligently adapts video format parameters of the input video before encoding, subsequently employing a CNN at the…
DecoderVideo CompressionA CNN-based Post-Processor for Perceptually-Optimized Immersive Media Compression
In recent years, resolution adaptation based on deep neural networks has enabled significant performance gains for conventional (2D) video codecs. This paper investigates the effectiveness of spatial resolution resamplin…
How Asynchronous Events Encode Video
As event-based sensing gains in popularity, theoretical understanding is needed to harness this technology's potential. Instead of recording video by capturing frames, event-based cameras have sensors that emit events wh…
Event-based visionDecomposition, Compression, and Synthesis (DCS)-based Video Coding: A Neural Exploration via Resolution-Adaptive Learning
Inspired by the facts that retinal cells actually segregate the visual scene into different attributes (e.g., spatial details, temporal motion) for respective neuronal processing, we propose to first decompose the input …
Motion CompensationSuper-ResolutionVideo CompressionVideo ReconstructionConvolutional sparse coding for capturing high speed video content
Video capture is limited by the trade-off between spatial and temporal resolution: when capturing videos of high temporal resolution, the spatial resolution decreases due to bandwidth limitations in the capture system. A…
Compressive SensingVocal Bursts Intensity Prediction