paper-with-me

홈 › Papers

JPD-SE: High-Level Semantics for Joint Perception-Distortion Enhancement in Image Compression

2020-05-24 · Shiyu Duan, Huaijin Chen, Jinwei Gu

While humans can effortlessly transform complex visual scenes into simple words and the other way around by leveraging their high-level understanding of the content, conventional or the more recent learned image compression codecs do not seem to utilize the semantic meanings of visual content to their full potential. Moreover, they focus mostly on rate-distortion and tend to underperform in perception quality especially in low bitrate regime, and often disregard the performance of downstream computer vision algorithms, which is a fast-growing consumer group of compressed images in addition to human viewers. In this paper, we (1) present a generic framework that can enable any image codec to leverage high-level semantics and (2) study the joint optimization of perception quality and distortion. Our idea is that given any codec, we utilize high-level semantics to augment the low-level visual features extracted by it and produce essentially a new, semantic-aware codec. We propose a three-phase training scheme that teaches semantic-aware codecs to leverage the power of semantic to jointly optimize rate-perception-distortion (R-PD) performance. As an additional benefit, semantic-aware codecs also boost the performance of downstream computer vision algorithms. To validate our claim, we perform extensive empirical evaluations and provide both quantitative and qualitative results.

📄 PDF Abstract BibTeX arXiv:2005.12810

Code (1)

sensebrain/jpd-se 공식 구현 pytorch

Tasks

Image Compression

Similar Papers 제목 키워드 기반

Deep Joint Source-Channel Coding Based on Semantics of Pixels

2022-08-24 · Qizheng Sun, Caili Guo, Yang Yang, Jiujiu Chen 외

The semantic information of the image for intelligent tasks is hidden behind the pixels, and slight changes in the pixels will affect the performance of intelligent tasks. In order to preserve semantic information behind…

Mitigating Perception Bias: A Training-Free Approach to Enhance LMM for Image Quality Assessment

2024-11-19 · Siyi Pan, Baoliang Chen, Danni Huang, Hanwei Zhu 외

Despite the impressive performance of large multimodal models (LMMs) in high-level visual tasks, their capacity for image quality assessment (IQA) remains limited. One main reason is that LMMs are primarily trained for h…

Image CaptioningImage Quality Assessment

Hierarchical Fusion and Joint Aggregation: A Multi-Level Feature Representation Method for AIGC Image Quality Assessment

2025-07-23 · Linghe Meng, Jiarun Song arxiv

The quality assessment of AI-generated content (AIGC) faces multi-dimensional challenges, that span from low-level visual perception to high-level semantic understanding. Existing methods generally rely on single-level v…

Image Quality Assessment

Semantic Communication via Rate Distortion Perception Bottleneck

2024-05-16 · Zihe Zhao, Chunyue Wang

With the advancement of Artificial Intelligence (AI) technology, next-generation wireless communication network is facing unprecedented challenge. Semantic communication has become a novel solution to address such challe…

Image ReconstructionSemantic Communication

Rateless Stochastic Coding for Delay-Constrained Semantic Communication

2024-06-28 · Cheng Peng, Rulong Wang, Yong Xiao

We consider the problem of joint source-channel coding for semantic communication from a rateless perspective, the purpose of which is to settle the balance between reliability (distortion/perception) and effectiveness (…

DecoderPerceptual DistanceQuantizationSemantic Communication