paper-with-me

Papers

Conditional Neural Video Coding with Spatial-Temporal Super-Resolution

2024-01-25 · Henan Wang, Xiaohan Pan, Runsen Feng, Zongyu Guo, Zhibo Chen

This document is an expanded version of a one-page abstract originally presented at the 2024 Data Compression Conference. It describes our proposed method for the video track of the Challenge on Learned Image Compression (CLIC) 2024. Our scheme follows the typical hybrid coding framework with some novel techniques. Firstly, we adopt Spynet network to produce accurate motion vectors for motion estimation. Secondly, we introduce the context mining scheme with conditional frame coding to fully exploit the spatial-temporal information. As for the low target bitrates given by CLIC, we integrate spatial-temporal super-resolution modules to improve rate-distortion performance. Our team name is IMCLVC.

📄 PDF Abstract BibTeX arXiv:2401.13959

Code (0)

등록된 구현이 없습니다.

Tasks

Data CompressionImage CompressionMotion EstimationSuper-Resolution

Similar Papers 제목 키워드 기반

On the Rate-Distortion-Complexity Trade-offs of Neural Video Coding

2024-10-04 · Yi-Hsin Chen, Kuan-Wei Ho, Martin Benjak, Jörn Ostermann 외

This paper aims to delve into the rate-distortion-complexity trade-offs of modern neural video coding. Recent years have witnessed much research effort being focused on exploring the full potential of neural video coding…

Learned Wavelet Video Coding using Motion Compensated Temporal Filtering

2023-05-25 · Anna Meyer, Fabian Brand, André Kaup

We present an end-to-end trainable wavelet video coder based on motion-compensated temporal filtering (MCTF). Thereby, we introduce a different coding scheme for learned video compression, which is currently dominated by…

Video Compression

MagicDriveDiT: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

2024-11-21 · Ruiyuan Gao, Kai Chen, Bo Xiao, Lanqing Hong 외

The rapid advancement of diffusion models has greatly improved video synthesis, especially in controllable video generation, which is essential for applications like autonomous driving. However, existing methods are limi…

Autonomous DrivingVideo Generation

Transcoded Video Restoration by Temporal Spatial Auxiliary Network

2021-12-15 · Li Xu, Gang He, Jinjia Zhou, Jie Lei 외

In most video platforms, such as Youtube, and TikTok, the played videos usually have undergone multiple video encodings such as hardware encoding by recording devices, software encoding by video editing apps, and single/…

Video EditingVideo Restoration

DeRA: Decoupled Representation Alignment for Video Tokenization

2025-12-04 · Pengbo Guo, Junke Wang, Zhen Xing, Chengxu Liu 외 arxiv

This paper presents DeRA, a novel 1D video tokenizer that decouples the spatial-temporal representation learning in video tokenization to achieve better training efficiency and performance. Specifically, DeRA maintains a…

Representation LearningVideo Generation