paper-with-me

Papers

Generative Video Motion Editing with 3D Point Tracks

2025-12-01 · Yao-Chih Lee, Zhoutong Zhang, Jiahui Huang, Jui-Hsien Wang, Joon-Young Lee, Jia-Bin Huang, Eli Shechtman, Zhengqi Li arxiv

Camera and object motions are central to a video's narrative. However, precisely editing these captured motions remains a significant challenge, especially under complex object movements. Current motion-controlled image-to-video (I2V) approaches often lack full-scene context for consistent video editing, while video-to-video (V2V) methods provide viewpoint changes or basic object translation, but offer limited control over fine-grained object motion. We present a track-conditioned V2V framework that enables joint editing of camera and object motion. We achieve this by conditioning a video generation model on a source video and paired 3D point tracks representing source and target motions. These 3D tracks establish sparse correspondences that transfer rich context from the source video to new motions while preserving spatiotemporal coherence. Crucially, compared to 2D tracks, 3D tracks provide explicit depth cues, allowing the model to resolve depth order and handle occlusions for precise motion editing. Trained in two stages on synthetic and real data, our model supports diverse motion edits, including joint camera/object manipulation, motion transfer, and non-rigid deformation, unlocking new creative potential in video editing.

📄 PDF Abstract BibTeX arXiv:2512.02015

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

Direct Motion Models for Assessing Generated Videos

2025-04-30 · Kelsey Allen, Carl Doersch, Guangyao Zhou, Mohammed Suhail 외

A current limitation of video generative video models is that they generate plausible looking frames, but poor motion -- an issue that is not well captured by FVD and other popular methods for evaluating generated videos…

Action Recognition

VideoHandles: Editing 3D Object Compositions in Videos Using Video Generative Priors

2025-03-03 · CVPR 2025 1 · Juil Koo, Paul Guerrero, Chun-Hao Paul Huang, Duygu Ceylan 외

Generative methods for image and video editing use generative models as priors to perform edits despite incomplete information, such as changing the composition of 3D objects shown in a single image. Recent methods have …

3D ReconstructionObjectVideo Editing

MotionV2V: Editing Motion in a Video

2025-11-25 · Ryan Burgert, Charles Herrmann, Forrester Cole, Michael S Ryoo 외 arxiv

While generative video models have achieved remarkable fidelity and consistency, applying these capabilities to video editing remains a complex challenge. Recent research has explored motion controllability as a means to…

Text-to-Video Generation

Point-to-Point: Sparse Motion Guidance for Controllable Video Editing

2025-11-23 · Yeji Song, Jaehyun Lee, Mijin Koo, JunHoo Lee 외 arxiv

Accurately preserving motion while editing a subject remains a core challenge in video editing tasks. Existing methods often face a trade-off between edit and motion fidelity, as they rely on motion representations that …

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis

2026-07-19 · Lingwei Dang, Juntong Li, Zonghan Li, Hongwen Zhang 외 hf

Hand-Object Interaction (HOI) synthesis is a cornerstone for animation production and embodied AI. Despite the strong priors of video foundation models, multi-view consistent HOI synthesis remains challenging due to comp…