paper-with-me

Papers

Learning Spatio-Temporal Downsampling for Effective Video Upscaling

2022-03-15 · Xiaoyu Xiang, Yapeng Tian, Vijay Rengarajan, Lucas Young, Bo Zhu, Rakesh Ranjan

Downsampling is one of the most basic image processing operations. Improper spatio-temporal downsampling applied on videos can cause aliasing issues such as moir\'e patterns in space and the wagon-wheel effect in time. Consequently, the inverse task of upscaling a low-resolution, low frame-rate video in space and time becomes a challenging ill-posed problem due to information loss and aliasing artifacts. In this paper, we aim to solve the space-time aliasing problem by learning a spatio-temporal downsampler. Towards this goal, we propose a neural network framework that jointly learns spatio-temporal downsampling and upsampling. It enables the downsampler to retain the key patterns of the original video and maximizes the reconstruction performance of the upsampler. To make the downsamping results compatible with popular image and video storage formats, the downsampling results are encoded to uint8 with a differentiable quantization layer. To fully utilize the space-time correspondences, we propose two novel modules for explicit temporal propagation and space-time feature rearrangement. Experimental results show that our proposed method significantly boosts the space-time reconstruction quality by preserving spatial textures and motion patterns in both downsampling and upscaling. Moreover, our framework enables a variety of applications, including arbitrary video resampling, blurry frame reconstruction, and efficient video storage.

📄 PDF Abstract BibTeX arXiv:2203.08140

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Similar Papers 제목 키워드 기반

Video Rescaling Networks with Joint Optimization Strategies for Downscaling and Upscaling

2021-03-27 · CVPR 2021 1 · Yan-Cheng Huang, Yi-Hsin Chen, Cheng-You Lu, Hui-Po Wang 외

This paper addresses the video rescaling task, which arises from the needs of adapting the video spatial resolution to suit individual viewing devices. We aim to jointly optimize video downscaling and upscaling as a comb…

Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models

2024-12-24 · CVPR 2025 1 · Jinhui Yi, Syed Talal Wasim, Yanan Luo, Muzammal Naseer 외

We present an efficient encoder-free approach for video-language understanding that achieves competitive performance while significantly reducing computational overhead. Current video-language models typically rely on he…

Question AnsweringVideo Question Answering

VoLUT: Efficient Volumetric streaming enhanced by LUT-based super-resolution

2025-02-17 · Chendong Wang, Anlan Zhang, Yifan Yang, Lili Qiu 외

3D volumetric video provides immersive experience and is gaining traction in digital media. Despite its rising popularity, the streaming of volumetric video content poses significant challenges due to the high data bandw…

Super-Resolution

Continuous Space-Time Video Super-Resolution with 3D Fourier Fields

2025-09-30 · Alexander Becker, Julius Erbach, Dominik Narnhofer, Konrad Schindler arxiv

We introduce a novel formulation for continuous space-time video super-resolution. Instead of decoupling the representation of a video sequence into separate spatial and temporal components and relying on brittle, explic…

Space-time Video Super-resolution

MiVID: Multi-Strategic Self-Supervision for Video Frame Interpolation using Diffusion Model

2025-11-08 · Priyansh Srivastava, Romit Chatterjee, Abir Sen, Aradhana Behura 외 arxiv

Video Frame Interpolation (VFI) remains a cornerstone in video enhancement, enabling temporal upscaling for tasks like slow-motion rendering, frame rate conversion, and video restoration. While classical methods rely on …

Video Frame InterpolationVideo RestorationVideo Enhancement