paper-with-me

홈 › Papers

Hierarchical Flow Diffusion for Efficient Frame Interpolation

2025-01-01 · CVPR 2025 1 · Yang Hai, Guo Wang, Tan Su, Wenjie Jiang, Yinlin Hu

Most recent diffusion-based methods still show a large gap compared to non-diffusion methods for video frame interpolation, in both accuracy and efficiency. Most of them formulate the problem as a denoising procedure in latent space directly, which is less effective caused by the large latent space. We propose to model bilateral optical flow explicitly by hierarchical diffusion models, which has much smaller search space in the denoising procedure. Based on the flow diffusion model, we then use a flow-guided image synthesizer to produce the final result. We train the flow diffusion model and the image synthesizer end to end. Our method achieves state of the art in accuracy, and 10+ times faster than other diffusion-based methods. The project page is at: https://hfd-interpolation.github.io.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingOptical Flow EstimationVideo Frame Interpolation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

SNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolation

2026-08-13 · Jisoo Jeong, Hong Cai, Jamie Menjay Lin, Hanno Ackermann 외 arxiv

We propose Symmetric Nonlinear Motion-guided Generative Video Frame Interpolation (SNM-VFI), a training-free framework for motion-controllable generative video frame interpolation with pre-trained optical flow and video …

Video Frame Interpolation

TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

2024-10-05 · Haiyang Liu, Xingchao Yang, Tomoya Akiyama, Yuantian Huang 외

We present TANGO, a framework for generating co-speech body-gesture videos. Given a few-minute, single-speaker reference video and target speech audio, TANGO produces high-fidelity videos with synchronized body gestures.…

cross-modal alignmentRetrievalvalid

Depth-Aware Video Frame Interpolation

2019-04-01 · CVPR 2019 6 · Wenbo Bao, Wei-Sheng Lai, Chao Ma, Xiaoyun Zhang 외

Video frame interpolation aims to synthesize nonexistent frames in-between the original frames. While significant advances have been made from the recent deep convolutional neural networks, the quality of interpolation i…

Optical Flow EstimationVideo Frame Interpolation

EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation

2025-03-20 · CVPR 2025 1 · Zihao Zhang, Haoran Chen, Haoyu Zhao, Guansong Lu 외

Handling complex or nonlinear motion patterns has long posed challenges for video frame interpolation. Although recent advances in diffusion-based methods offer improvements over traditional optical flow-based approaches…

Optical Flow EstimationVideo Frame Interpolation

RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation

2026-04-21 · Ahmed Marouane Djouamaa, Abir Belaala, Abdellah Zakaria Sellam, Salah Eddine Bekhouche 외 arxiv

Accurate medical image segmentation requires both long-range contextual reasoning and precise boundary delineation, a task where existing transformer- and diffusion-based paradigms are frequently bottlenecked by quadrati…

Medical Image Segmentation