FD-LSCIC: Frequency Decomposition-based Learned Screen Content Image Compression
The learned image compression (LIC) methods have already surpassed traditional techniques in compressing natural scene (NS) images. However, directly applying these methods to screen content (SC) images, which possess distinct characteristics such as sharp edges, repetitive patterns, embedded text and graphics, yields suboptimal results. This paper addresses three key challenges in SC image compression: learning compact latent features, adapting quantization step sizes, and the lack of large SC datasets. To overcome these challenges, we propose a novel compression method that employs a multi-frequency two-stage octave residual block (MToRB) for feature extraction, a cascaded triple-scale feature fusion residual block (CTSFRB) for multi-scale feature integration and a multi-frequency context interaction module (MFCIM) to reduce inter-frequency correlations. Additionally, we introduce an adaptive quantization module that learns scaled uniform noise for each frequency component, enabling flexible control over quantization granularity. Furthermore, we construct a large SC image compression dataset (SDU-SCICD10K), which includes over 10,000 images spanning basic SC images, computer-rendered images, and mixed NS and SC images from both PC and mobile platforms. Experimental results demonstrate that our approach significantly improves SC image compression performance, outperforming traditional standards and state-of-the-art learning-based methods in terms of peak signal-to-noise ratio (PSNR) and multi-scale structural similarity (MS-SSIM).
Code (0)
등록된 구현이 없습니다.
Tasks
Image CompressionMS-SSIMQuantizationSSIMMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
NeR-SC: Adapting Neural Video Representation to Screen Content
Implicit neural representations have emerged as a promising paradigm for video compression, with recent methods achieving competitive performance on natural video. However, screen content video -- common in remote deskto…
Screen Content Image Segmentation Using Robust Regression and Sparse Decomposition
This paper considers how to separate text and/or graphics from smooth background in screen content and mixed document images and proposes two approaches to perform this segmentation task. The proposed methods make use of…
Image SegmentationMedical Image SegmentationregressionSemantic SegmentationMoireé Pattern Detection using Wavelet Decomposition and Convolutional Neural Network
Moiré patterns are interference patterns that are produced due to the overlap of the digital grids of the camera sensor resulting in a high-frequency noise in the image. This paper proposes a new method to detect Moiré p…
Towards General Game Representations: Decomposing Games Pixels into Content and Style
On-screen game footage contains rich contextual information that players process when playing and experiencing a game. Learning pixel representations of games can benefit artificial intelligence across several downstream…
OMR-NET: a two-stage octave multi-scale residual network for screen content image compression
Screen content (SC) differs from natural scene (NS) with unique characteristics such as noise-free, repetitive patterns, and high contrast. Aiming at addressing the inadequacies of current learned image compression (LIC)…
Image Compression