paper-with-me

Papers

Active Sampling for Ultra-Low-Bit-Rate Video Compression via Conditional Controlled Diffusion

2026-05-04 · Amirhosein Javadi, Shirin Saeedi Bidokhti, Tara Javidi arxiv

Diffusion models provide a powerful generative prior for perceptual reconstruction at ultra-low bitrates, but effective video compression requires controlling the generative process using highly compact conditioning signals. In this work, we present ActDiff-VC, a diffusion-based video compression framework for the ultra-low-bitrate regime. Our method partitions videos into variable-length segments, transmits keyframes only when needed, and summarizes temporal dynamics using a compact set of tracked point trajectories. Conditioned on these sparse signals, a conditional diffusion decoder synthesizes the remaining frames, enabling perceptually realistic reconstruction under severe rate constraints. To support this design, we introduce two mechanisms: content-adaptive keyframe selection and budget-aware sparse trajectory selection, which together enable compact yet effective conditioning for generative reconstruction. Experiments on the UVG and MCL-JCV benchmarks show that ActDiff-VC achieves up to 64.6\% bitrate reduction at matched NIQE, improves KID by up to 64.6\% and FID by up to 37.7\% at comparable bitrates against strong learned codecs, and delivers favorable perceptual rate--distortion trade-offs relative to learned and diffusion-based baselines in the ultra-low-bitrate regime.

📄 PDF Abstract BibTeX arXiv:2605.02849

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Compressing Human Body Video with Interactive Semantics: A Generative Approach

2025-05-22 · Bolin Chen, Shanzhi Yin, Hanwei Zhu, Lingyu Zhu 외

In this paper, we propose to compress human body video with interactive semantics, which can facilitate video coding to be interactive and controllable by manipulating semantic-level representations embedded in the coded…

DecoderVideo Reconstruction

Deep Sylvester Posterior Inference for Adaptive Compressed Sensing in Ultrasound Imaging

2025-01-07 · Simon W. Penninga, Hans van Gorp, Ruud J. G. van Sloun

Ultrasound images are commonly formed by sequential acquisition of beam-steered scan-lines. Minimizing the number of required scan-lines can significantly enhance frame rate, field of view, energy efficiency, and data tr…

compressed sensing

TVC: Tokenized Video Compression with Ultra-Low Bitrate

2025-04-22 · Lebin Zhou, Cihan Ruan, Nam Ling, Wei Wang 외

Tokenized visual representations have shown great promise in image compression, yet their extension to video remains underexplored due to the challenges posed by complex temporal dynamics and stringent bitrate constraint…

DecoderImage CompressionVideo Compression

Generative Latent Coding for Ultra-Low Bitrate Image and Video Compression

2025-05-22 · Linfeng Qi, Zhaoyang Jia, Jiahao Li, Bin Li 외

Most existing approaches for image and video compression perform transform coding in the pixel space to reduce redundancy. However, due to the misalignment between the pixel-space distortion and human perception, such sc…

Image CompressionVideo Compression

Block Modulating Video Compression: An Ultra Low Complexity Image Compression Encoder for Resource Limited Platforms

2022-05-07 · Siming Zheng, Yujia Xue, Waleed Tahir, Zhengjue Wang 외

We consider the image and video compression on resource limited platforms. An ultra low-cost image encoder, named Block Modulating Video Compression (BMVC) with an encoding complexity ${\cal O}(1)$ is proposed to be impl…

DecoderImage CompressionQuantizationVideo Compression