paper-with-me

Papers

Controllable Human-centric Keyframe Interpolation with Generative Prior

2025-06-03 · Zujin Guo, Size Wu, Zhongang Cai, Wei Li, Chen Change Loy

Existing interpolation methods use pre-trained video diffusion priors to generate intermediate frames between sparsely sampled keyframes. In the absence of 3D geometric guidance, these methods struggle to produce plausible results for complex, articulated human motions and offer limited control over the synthesized dynamics. In this paper, we introduce PoseFuse3D Keyframe Interpolator (PoseFuse3D-KI), a novel framework that integrates 3D human guidance signals into the diffusion process for Controllable Human-centric Keyframe Interpolation (CHKI). To provide rich spatial and structural cues for interpolation, our PoseFuse3D, a 3D-informed control model, features a novel SMPL-X encoder that transforms 3D geometry and shape into the 2D latent conditioning space, alongside a fusion network that integrates these 3D cues with 2D pose embeddings. For evaluation, we build CHKI-Video, a new dataset annotated with both 2D poses and 3D SMPL-X parameters. We show that PoseFuse3D-KI consistently outperforms state-of-the-art baselines on CHKI-Video, achieving a 9% improvement in PSNR and a 38% reduction in LPIPS. Comprehensive ablations demonstrate that our PoseFuse3D model improves interpolation fidelity.

📄 PDF Abstract BibTeX arXiv:2506.03119

Code (0)

등록된 구현이 없습니다.

Tasks

3D geometry

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

I2V3D: Controllable image-to-video generation with 3D guidance

2025-03-12 · Zhiyuan Zhang, Dongdong Chen, Jing Liao

We present I2V3D, a novel framework for animating static images into dynamic videos with precise 3D control, leveraging the strengths of both 3D geometry guidance and advanced generative models. Our approach combines the…

3D geometryImage to Video GenerationVideo Generation

Controllable Generative Video Compression

2026-04-08 · Ding Ding, Daowen Li, Ying Chen, Yixin Gao 외 arxiv

Perceptual video compression adopts generative video modeling to improve perceptual realism but frequently sacrifices signal fidelity, diverging from the goal of video compression to faithfully reproduce visual signal. T…

Video Generation

Generative Inbetweening: Adapting Image-to-Video Models for Keyframe Interpolation

2024-08-27 · Xiaojuan Wang, Boyang Zhou, Brian Curless, Ira Kemelmacher-Shlizerman 외

We present a method for generating video sequences with coherent motion between a pair of input key frames. We adapt a pretrained large-scale image-to-video diffusion model (originally trained to generate videos moving f…

M3-CVC: Controllable Video Compression with Multimodal Generative Models

2024-11-24 · Rui Wan, Qi Zheng, Yibo Fan

Traditional and neural video codecs commonly encounter limitations in controllability and generality under ultra-low-bitrate coding scenarios. To overcome these challenges, we propose M3-CVC, a controllable video compres…

Video Compression

Topology Aware Neural Interpolation of Scalar Fields

2025-08-25 · Mohamed Kissi, Keanu Sisouk, Joshua A. Levine, Julien Tierny arxiv

This paper presents a neural scheme for the topology-aware interpolation of time-varying scalar fields. Given a time-varying sequence of persistence diagrams, along with a sparse temporal sampling of the corresponding sc…