Exploring Effective Mask Sampling Modeling for Neural Image Compression
Image compression aims to reduce the information redundancy in images. Most existing neural image compression methods rely on side information from hyperprior or context models to eliminate spatial redundancy, but rarely address the channel redundancy. Inspired by the mask sampling modeling in recent self-supervised learning methods for natural language processing and high-level vision, we propose a novel pretraining strategy for neural image compression. Specifically, Cube Mask Sampling Module (CMSM) is proposed to apply both spatial and channel mask sampling modeling to image compression in the pre-training stage. Moreover, to further reduce channel redundancy, we propose the Learnable Channel Mask Module (LCMM) and the Learnable Channel Completion Module (LCCM). Our plug-and-play CMSM, LCMM, LCCM modules can apply to both CNN-based and Transformer-based architectures, significantly reduce the computational cost, and improve the quality of images. Experiments on the public Kodak and Tecnick datasets demonstrate that our method achieves competitive performance with lower computational complexity compared to state-of-the-art image compression methods.
Code (0)
등록된 구현이 없습니다.
Tasks
Image CompressionSelf-Supervised LearningSimilar Papers 제목 키워드 기반
Exploring the Coordination of Frequency and Attention in Masked Image Modeling
Recently, masked image modeling (MIM), which learns visual representations by reconstructing the masked patches of an image, has dominated self-supervised learning in computer vision. However, the pre-training of MIM alw…
AttributeRepresentation LearningSelf-Supervised LearningAdaptive Mask Sampling and Manifold to Euclidean Subspace Learning with Distance Covariance Representation for Hyperspectral Image Classification
For the abundant spectral and spatial information recorded in hyperspectral images (HSIs), fully exploring spectral-spatial relationships has attracted widespread attention in hyperspectral image classification (HSIC) co…
Hyperspectral image analysisHyperspectral Image ClassificationHyperspectral Image Segmentationimage-classification+1Improving Masked Autoencoders by Learning Where to Mask
Masked image modeling is a promising self-supervised learning method for visual data. It is typically built upon image patches with random masks, which largely ignores the variation of information density between them. T…
Image ReconstructionSelf-Supervised LearningExposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings
While recent work has shown that scores from models trained by the ubiquitous masked language modeling (MLM) objective effectively discriminate probable from improbable sequences, it is still an open question if these ML…
Language ModelingLanguage ModellingMachine TranslationMasked Language Modeling+2Exploring Masked Autoencoders for Sensor-Agnostic Image Retrieval in Remote Sensing
Self-supervised learning through masked autoencoders (MAEs) has recently attracted great attention for remote sensing (RS) image representation learning, and thus embodies a significant potential for content-based image …
Content-Based Image RetrievalImage RetrievalRepresentation LearningRetrieval+1