paper-with-me

홈 › Papers

LADDER: An Efficient Framework for Video Frame Interpolation

2024-04-17 · Tong Shen, Dong Li, Ziheng Gao, Lu Tian, Emad Barsoum

Video Frame Interpolation (VFI) is a crucial technique in various applications such as slow-motion generation, frame rate conversion, video frame restoration etc. This paper introduces an efficient video frame interpolation framework that aims to strike a favorable balance between efficiency and quality. Our framework follows a general paradigm consisting of a flow estimator and a refinement module, while incorporating carefully designed components. First of all, we adopt depth-wise convolution with large kernels in the flow estimator that simultaneously reduces the parameters and enhances the receptive field for encoding rich context and handling complex motion. Secondly, diverging from a common design for the refinement module with a UNet-structure (encoder-decoder structure), which we find redundant, our decoder-only refinement module directly enhances the result from coarse to fine features, offering a more efficient process. In addition, to address the challenge of handling high-definition frames, we also introduce an innovative HD-aware augmentation strategy during training, leading to consistent enhancement on HD images. Extensive experiments are conducted on diverse datasets, Vimeo90K, UCF101, Xiph and SNU-FILM. The results demonstrate that our approach achieves state-of-the-art performance with clear improvement while requiring much less FLOPs and parameters, reaching to a better spot for balancing efficiency and quality.

📄 PDF Abstract BibTeX arXiv:2404.11108

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderMotion GenerationVideo Frame Interpolation

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Efficient Bitrate Ladder Construction for Content-Optimized Adaptive Video Streaming

2021-02-08

One of the challenges faced by many video providers is the heterogeneity of network specifications, user requirements, and content compression performance. The universal solution of a fixed bitrate ladder is inadequate i…

LaDDer: Latent Data Distribution Modelling with a Generative Prior

2020-08-31 · Shuyu Lin, Ronald Clark

In this paper, we show that the performance of a learnt generative model is closely related to the model's ability to accurately represent the inferred \textbf{latent data distribution}, i.e. its topology and structural …

Representation Learning

VMAF-based Bitrate Ladder Estimation for Adaptive Streaming

2021-03-12 · Angeliki V. Katsenou, Fan Zhang, Kyle Swanson, Mariana Afonso 외

In HTTP Adaptive Streaming, video content is conventionally encoded by adapting its spatial resolution and quantization level to best match the prevailing network state and display characteristics. It is well known that …

Quantization

Leveraging Compression to Construct Transferable Bitrate Ladders

2025-12-15 · Krishna Srikar Durbha, Hassene Tmar, Ping-Hao Wu, Ioannis Katsavounidis 외 arxiv

Over the past few years, per-title and per-shot video encoding techniques have demonstrated significant gains as compared to conventional techniques such as constant CRF encoding and the fixed bitrate ladder. These techn…

Bitrate Ladder Construction using Visual Information Fidelity

2023-12-12 · Krishna Srikar Durbha, Hassene Tmar, Cosmin Stejerean, Ioannis Katsavounidis 외

Recently proposed perceptually optimized per-title video encoding methods provide better BD-rate savings than fixed bitrate-ladder approaches that have been employed in the past. However, a disadvantage of per-title enco…