paper-with-me

Papers

Controllable Generative Video Compression

2026-04-08 · Ding Ding, Daowen Li, Ying Chen, Yixin Gao, Ruixiao Dong, Kai Li, Li Li arxiv

Perceptual video compression adopts generative video modeling to improve perceptual realism but frequently sacrifices signal fidelity, diverging from the goal of video compression to faithfully reproduce visual signal. To alleviate the dilemma between perception and fidelity, in this paper we propose Controllable Generative Video Compression (CGVC) paradigm to faithfully generate details guided by multiple visual conditions. Under the paradigm, representative keyframes of the scene are coded and used to provide structural priors for non-keyframe generation. Dense per-frame control prior is additionally coded to better preserve finer structure and semantics of each non-keyframe. Guided by these priors, non-keyframes are reconstructed by controllable video generation model with temporal and content consistency. Furthermore, to accurately recover color information of the video, we develop a color-distance-guided keyframe selection algorithm to adaptively choose keyframes. Experimental results show CGVC outperforms previous perceptual video compression method in terms of both signal fidelity and perceptual quality.

📄 PDF Abstract BibTeX arXiv:2604.06655

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

M3-CVC: Controllable Video Compression with Multimodal Generative Models

2024-11-24 · Rui Wan, Qi Zheng, Yibo Fan

Traditional and neural video codecs commonly encounter limitations in controllability and generality under ultra-low-bitrate coding scenarios. To overcome these challenges, we propose M3-CVC, a controllable video compres…

Video Compression

Compressing Human Body Video with Interactive Semantics: A Generative Approach

2025-05-22 · Bolin Chen, Shanzhi Yin, Hanwei Zhu, Lingyu Zhu 외

In this paper, we propose to compress human body video with interactive semantics, which can facilitate video coding to be interactive and controllable by manipulating semantic-level representations embedded in the coded…

DecoderVideo Reconstruction

Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis

2025-05-29 · Hengyuan Cao, Yutong Feng, Biao Gong, Yijing Tian 외

Video generative models can be regarded as world simulators due to their ability to capture dynamic, continuous changes inherent in real-world environments. These models integrate high-dimensional information across visu…

Dimensionality ReductionImage Generation

Interactive Face Video Coding: A Generative Compression Framework

2023-02-20 · Bolin Chen, Zhao Wang, Binzhe Li, Shurun Wang 외

In this paper, we propose a novel framework for Interactive Face Video Coding (IFVC), which allows humans to interact with the intrinsic visual representations instead of the signals. The proposed solution enjoys several…

Once-for-All: Controllable Generative Image Compression with Dynamic Granularity Adaption

2024-06-02 · Anqi Li, Feng Li, Yuxi Liu, Runmin Cong 외

Although recent generative image compression methods have demonstrated impressive potential in optimizing the rate-distortion-perception trade-off, they still face the critical challenge of flexible rate adaption to dive…

AllImage Compression