paper-with-me

Papers

Multi-Scale Deformable Alignment and Content-Adaptive Inference for Flexible-Rate Bi-Directional Video Compression

2023-06-28 · M. Akin Yilmaz, O. Ugur Ulas, A. Murat Tekalp

The lack of ability to adapt the motion compensation model to video content is an important limitation of current end-to-end learned video compression models. This paper advances the state-of-the-art by proposing an adaptive motion-compensation model for end-to-end rate-distortion optimized hierarchical bi-directional video compression. In particular, we propose two novelties: i) a multi-scale deformable alignment scheme at the feature level combined with multi-scale conditional coding, ii) motion-content adaptive inference. In addition, we employ a gain unit, which enables a single model to operate at multiple rate-distortion operating points. We also exploit the gain unit to control bit allocation among intra-coded vs. bi-directionally coded frames by fine tuning corresponding models for truly flexible-rate learned video coding. Experimental results demonstrate state-of-the-art rate-distortion performance exceeding those of all prior art in learned video coding.

📄 PDF Abstract BibTeX arXiv:2306.16544

Code (1)

KUIS-AI-Tekalp-Research-Group/video-compression 공식 구현 pytorch

Tasks

Motion CompensationVideo Compression

Similar Papers 제목 키워드 기반

A Universal Scale-Adaptive Deformable Transformer for Image Restoration across Diverse Artifacts

2025-01-01 · CVPR 2025 1 · Xuyi He, Yuhui Quan, Ruotao Xu, Hui Ji

Structured artifacts are semi-regular, repetitive patterns that closely intertwine with genuine image content, making their removal highly challenging. In this paper, we introduce the Scale-Adaptive Deformable Transf…

Image RestorationRain Removal

Mixture of Scale Experts for Alignment-free RGBT Video Object Detection and A Unified Benchmark

2024-10-16 · Qishun Wang, Zhengzheng Tu, Kunpeng Wang, Le Gu 외

Existing RGB-Thermal Video Object Detection (RGBT VOD) methods predominantly rely on the manual alignment of image pairs, that is both labor-intensive and time-consuming. This dependency significantly restricts the scala…

object-detectionObject DetectionVideo Object Detection

Content Adaptive based Motion Alignment Framework for Learned Video Compression

2025-12-15 · Tiange Zhang, Xiandong Meng, Siwei Ma arxiv

Recent advances in end-to-end video compression have shown promising results owing to their unified end-to-end learning optimization. However, such generalized frameworks often lack content-specific adaptation, leading t…

Learned Video Compression via Heterogeneous Deformable Compensation Network

2022-07-11 · Huairui Wang, Zhenzhong Chen, Chang Wen Chen

Learned video compression has recently emerged as an essential research topic in developing advanced video compression technologies, where motion compensation is considered one of the most challenging issues. In this pap…

Motion CompensationOptical Flow EstimationVideo Compression

GeoMAD: Geometry-Aware Multi-View Anomaly Detection via Deformable Fusion and Distributional Alignment

2026-08-27 · Shang-Fu Chen, Jhih-Ciang Wu, Kuan-Chuan Peng, Wen-Huang Cheng 외 arxiv

Multi-view anomaly detection (MvAD) detects defects by exploiting complementary observations from multiple camera viewpoints. The central challenge is to fuse views with sufficient geometric awareness while remaining sca…

Anomaly Detection