paper-with-me

홈 › Papers

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting

2025-06-05 · Yue Ma, Kunyu Feng, Xinhua Zhang, Hongyu Liu, David Junhao Zhang, Jinbo Xing, Yinhan Zhang, Ayden Yang, Zeyu Wang, Qifeng Chen

We introduce Follow-Your-Creation, a novel 4D video creation framework capable of both generating and editing 4D content from a single monocular video input. By leveraging a powerful video inpainting foundation model as a generative prior, we reformulate 4D video creation as a video inpainting task, enabling the model to fill in missing content caused by camera trajectory changes or user edits. To facilitate this, we generate composite masked inpainting video data to effectively fine-tune the model for 4D video generation. Given an input video and its associated camera trajectory, we first perform depth-based point cloud rendering to obtain invisibility masks that indicate the regions that should be completed. Simultaneously, editing masks are introduced to specify user-defined modifications, and these are combined with the invisibility masks to create a composite masks dataset. During training, we randomly sample different types of masks to construct diverse and challenging inpainting scenarios, enhancing the model's generalization and robustness in various 4D editing and generation tasks. To handle temporal consistency under large camera motion, we design a self-iterative tuning strategy that gradually increases the viewing angles during training, where the model is used to generate the next-stage training data after each fine-tuning iteration. Moreover, we introduce a temporal packaging module during inference to enhance generation quality. Our method effectively leverages the prior knowledge of the base model without degrading its original performance, enabling the generation of 4D videos with consistent multi-view coherence. In addition, our approach supports prompt-based content editing, demonstrating strong flexibility and significantly outperforming state-of-the-art methods in both quality and versatility.

📄 PDF Abstract BibTeX arXiv:2506.04590

Code (0)

등록된 구현이 없습니다.

Tasks

Video GenerationVideo Inpainting

Methods 이 논문이 사용한 방법론

Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.
BASE 설명 없음

Similar Papers 제목 키워드 기반

CharacterChat: Supporting the Creation of Fictional Characters through Conversation and Progressive Manifestation with a Chatbot

2021-06-23 · Oliver Schmitt, Daniel Buschek

We present CharacterChat, a concept and chatbot to support writers in creating fictional characters. Concretely, writers progressively turn the bot into their imagined character through conversation. We iteratively devel…

ChatbotLanguage ModelingLanguage Modelling

See Your Heart: Psychological states Interpretation through Visual Creations

2023-02-11 · Likun Yang, Xiaokun Feng, Xiaotang Chen, Shiyu Zhang 외

In psychoanalysis, generating interpretations to one's psychological state through visual creations is facing significant demands. The two main tasks of existing studies in the field of computer vision, sentiment/emotion…

Emotion ClassificationImage Captioning

Expressive Communication: A Common Framework for Evaluating Developments in Generative Models and Steering Interfaces

2021-11-29 · Ryan Louie, Jesse Engel, Anna Huang

There is an increasing interest from ML and HCI communities in empowering creators with better generative models and more intuitive interfaces with which to control them. In music, ML researchers have focused on training…

Muizalix Pro

2025-02-07 · 02/07 2025 2 · Muizalix Pro

At Muizalix, we’re a dedicated team of creative professionals and digital experts passionate about helping businesses grow. Specializing in social media marketing, SEO, content creation, and website design, we provide co…

Marketing

A tool suite for creating question answering benchmarks

2014-05-01 · LREC 2014 5 · Axel-Cyrille Ngonga Ngomo, Norman Heino, Ren{\'e} Speck, Prodromos Malakasiotis

We introduce the BIOASQ suite, a set of open-source Web tools for the creation, assessment and community-driven improvement of question answering benchmarks. The suite comprises three main tools: (1) the annotation tool …

Question AnsweringRetrieval