paper-with-me

Papers

Video Motion Graphs

2025-03-26 · Haiyang Liu, Zhan Xu, Fa-Ting Hong, Hsin-Ping Huang, Yi Zhou, Yang Zhou

We present Video Motion Graphs, a system designed to generate realistic human motion videos. Using a reference video and conditional signals such as music or motion tags, the system synthesizes new videos by first retrieving video clips with gestures matching the conditions and then generating interpolation frames to seamlessly connect clip boundaries. The core of our approach is HMInterp, a robust Video Frame Interpolation (VFI) model that enables seamless interpolation of discontinuous frames, even for complex motion scenarios like dancing. HMInterp i) employs a dual-branch interpolation approach, combining a Motion Diffusion Model for human skeleton motion interpolation with a diffusion-based video frame interpolation model for final frame generation. ii) adopts condition progressive training to effectively leverage identity strong and weak conditions, such as images and pose. These designs ensure both high video texture quality and accurate motion trajectory. Results show that our Video Motion Graphs outperforms existing generative- and retrieval-based methods for multi-modal conditioned human motion video generation. Project page can be found at https://h-liu1997.github.io/Video-Motion-Graphs/

📄 PDF Abstract BibTeX arXiv:2503.20218

Code (0)

등록된 구현이 없습니다.

Tasks

Motion InterpolationVideo Frame InterpolationVideo Generation

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DreamLoop: Controllable Cinemagraph Generation from a Single Photograph

2026-01-06 · Aniruddha Mahapatra, Long Mai, Cusuh Ham, Feng Liu arxiv

Cinemagraphs, which combine static photographs with selective, looping motion, offer unique artistic appeal. Generating them from a single photograph in a controllable manner is particularly challenging. Existing image-a…

Text-Guided Synthesis of Eulerian Cinemagraphs

2023-07-06 · Aniruddha Mahapatra, Aliaksandr Siarohin, Hsin-Ying Lee, Sergey Tulyakov 외

We introduce Text2Cinemagraph, a fully automated method for creating cinemagraphs from text descriptions - an especially challenging task when prompts feature imaginary elements and artistic styles, given the complexity …

Image Animation

Video Captioning with Aggregated Features Based on Dual Graphs and Gated Fusion

2023-08-13 · Yutao Jin, Bin Liu, Jing Wang

The application of video captioning models aims at translating the content of videos by using accurate natural language. Due to the complex nature inbetween object interaction in the video, the comprehensive understandin…

Video Captioning

Motion Blender Gaussian Splatting for Dynamic Scene Reconstruction

2025-03-12 · Xinyu Zhang, Haonan Chang, YuHan Liu, Abdeslam Boularias

Gaussian splatting has emerged as a powerful tool for high-fidelity reconstruction of dynamic scenes. However, existing methods primarily rely on implicit motion representations, such as encoding motions into neural netw…

Dynamic ReconstructionSimulated Gaussian Manipulation

MovieGraphs: Towards Understanding Human-Centric Situations from Videos

2017-12-19 · CVPR 2018 6 · Paul Vicol, Makarand Tapaswi, Lluis Castrejon, Sanja Fidler

There is growing interest in artificial intelligence to build socially intelligent robots. This requires machines to have the ability to "read" people's emotions, motivations, and other factors that affect behavior. Towa…

Common Sense Reasoning