CAESR: Conditional Autoencoder and Super-Resolution for Learned Spatial Scalability
In this paper, we present CAESR, an hybrid learning-based coding approach for spatial scalability based on the versatile video coding (VVC) standard. Our framework considers a low-resolution signal encoded with VVC intra-mode as a base-layer (BL), and a deep conditional autoencoder with hyperprior (AE-HP) as an enhancement-layer (EL) model. The EL encoder takes as inputs both the upscaled BL reconstruction and the original image. Our approach relies on conditional coding that learns the optimal mixture of the source and the upscaled BL image, enabling better performance than residual coding. On the decoder side, a super-resolution (SR) module is used to recover high-resolution details and invert the conditional coding process. Experimental results have shown that our solution is competitive with the VVC full-resolution intra coding while being scalable.
Code (0)
등록된 구현이 없습니다.
Tasks
DecoderSuper-ResolutionSimilar Papers 제목 키워드 기반
Unsupervised Real Image Super-Resolution via Generative Variational AutoEncoder
Benefited from the deep learning, image Super-Resolution has been one of the most developing research fields in computer vision. Depending upon whether using a discriminator or not, a deep convolutional neural network ca…
DecoderDenoisingImage DenoisingImage Super-Resolution+1Patch-PODiff-ViT: Structured Latent Diffusion with Patchwise POD for Super-Resolution and Uncertainty Quantification
Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned latent spaces often lack interpretable uncertainty quantification. We i…
MaskCRT: Masked Conditional Residual Transformer for Learned Video Compression
Conditional coding has lately emerged as the mainstream approach to learned video compression. However, a recent study shows that it may perform worse than residual coding when the information bottleneck arises. Conditio…
MS-SSIMSSIMVideo CompressionTzK Flow - Conditional Generative Model
We introduce TzK (pronounced "task"), a conditional probability flow-based model that exploits attributes (e.g., style, class membership, or other side information) in order to learn tight conditional prior around manifo…
modelDiverse super-resolution with pretrained deep hiererarchical VAEs
We investigate the problem of producing diverse solutions to an image super-resolution problem. From a probabilistic perspective, this can be done by sampling from the posterior distribution of an inverse problem, which …
Computational EfficiencyImage Super-ResolutionSuper-Resolution