paper-with-me

홈 › Papers

CMC-Bench: Towards a New Paradigm of Visual Signal Compression

2024-06-13 · Chunyi Li, Xiele Wu, HaoNing Wu, Donghui Feng, ZiCheng Zhang, Guo Lu, Xiongkuo Min, Xiaohong Liu, Guangtao Zhai, Weisi Lin

Ultra-low bitrate image compression is a challenging and demanding topic. With the development of Large Multimodal Models (LMMs), a Cross Modality Compression (CMC) paradigm of Image-Text-Image has emerged. Compared with traditional codecs, this semantic-level compression can reduce image data size to 0.1\% or even lower, which has strong potential applications. However, CMC has certain defects in consistency with the original image and perceptual quality. To address this problem, we introduce CMC-Bench, a benchmark of the cooperative performance of Image-to-Text (I2T) and Text-to-Image (T2I) models for image compression. This benchmark covers 18,000 and 40,000 images respectively to verify 6 mainstream I2T and 12 T2I models, including 160,000 subjective preference scores annotated by human experts. At ultra-low bitrates, this paper proves that the combination of some I2T and T2I models has surpassed the most advanced visual signal codecs; meanwhile, it highlights where LMMs can be further optimized toward the compression task. We encourage LMM developers to participate in this test to promote the evolution of visual signal codec protocols.

📄 PDF Abstract BibTeX arXiv:2406.09356

Code (1)

q-future/cmc-bench 공식 구현 pytorch

Tasks

Image CompressionImage to text

Similar Papers 제목 키워드 기반

Controllable Generative Video Compression

2026-04-08 · Ding Ding, Daowen Li, Ying Chen, Yixin Gao 외 arxiv

Perceptual video compression adopts generative video modeling to improve perceptual realism but frequently sacrifices signal fidelity, diverging from the goal of video compression to faithfully reproduce visual signal. T…

Video Generation

Cross Modal Compression: Towards Human-comprehensible Semantic Compression

2022-09-06 · Jiguo Li, Chuanmin Jia, Xinfeng Zhang, Siwei Ma 외

Traditional image/video compression aims to reduce the transmission/storage cost with signal fidelity as high as possible. However, with the increasing demand for machine analysis and semantic monitoring in recent years,…

Feature CompressionSemantic CompressionVideo Compression

Rethinking Generative Human Video Coding with Implicit Motion Transformation

2025-06-12 · Bolin Chen, Ru-Ling Liao, Jie Chen, Yan Ye

Beyond traditional hybrid-based video codec, generative video codec could achieve promising compression performance by evolving high-dimensional signals into compact feature representations for bitstream compactness at t…

DecoderVideo Compression

Emerging Advances in Learned Video Compression: Models, Systems and Beyond

2025-04-30 · Chuanmin Jia, Feng Ye, Siwei Ma, Wen Gao 외

Video compression is a fundamental topic in the visual intelligence, bridging visual signal sensing/capturing and high-level visual analytics. The broad success of artificial intelligence (AI) technology has enriched the…

Video Compression

3DGS-VBench: A Comprehensive Video Quality Evaluation Benchmark for 3DGS Compression

2025-08-09 · Yuke Xing, William Gordon, Qi Yang, Kaifa Yang 외 arxiv

3D Gaussian Splatting (3DGS) enables real-time novel view synthesis with high visual fidelity, but its substantial storage requirements hinder practical deployment, prompting state-of-the-art (SOTA) 3DGS methods to incor…

Video Quality AssessmentNovel View Synthesis