paper-with-me

Papers

TeleMorpher: Toward Robust Simultaneous Motion-Location Editing

2026-06-18 · Haengbok Chung arxiv

Diffusion models have achieved remarkable success in image and video generation and editing. While recent studies have extended these efforts toward motion editing, simultaneously transforming both motion and location-despite its practical importance-remains largely unexplored. To better understand robust motion-location editing, we first analyze the fundamental factors that degrade its quality. Based on this analysis, we propose TeleMorpher, one of the first one-shot frameworks to the best of our knowledge, for simultaneous motion-location editing. Our approach leverages motion priors, a target motion-centric video generated from an off-the-shelf model as motion-editing guidance, and the ground truth motion to enable more controllable and precise motion-location editing. Via this, our framework works as follows: (1) we first disentangle the protagonist and the background via pre-trained segmentation and inpainting models. (2) Then, we introduce a training-free pose warping that edits the protagonist's motion with the motion prior as the guidance. (3) The result of warped motion video is directly injected into a baseline motion editor during inference, mitigating the difference between source and target motions while preserving the appearance of the source video. (4) To enhance the reliability of quantitative evaluations, we propose two new LPIPS-based metrics that measure the background consistency before and after the motion editing and the fidelity of motion editing performance via measuring the difference between the extracted protagonist's skeletons from source and target videos. Experiments with in-the-wild videos and the TaiChi dataset demonstrate that TeleMorpher achieves superior performance across both quantitative and qualitative measurements (real-human evaluation), underscoring its effectiveness.

📄 PDF Abstract BibTeX arXiv:2606.19676

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

ReVideo: Remake a Video with Motion and Content Control

2024-05-22 · Chong Mou, Mingdeng Cao, Xintao Wang, Zhaoyang Zhang 외

Despite significant advancements in video generation and editing using diffusion models, achieving accurate and localized video editing remains a substantial challenge. Additionally, most existing video editing methods p…

Video EditingVideo Generation

SAMA: Factorized Semantic Anchoring and Motion Alignment for Instruction-Guided Video Editing

2026-03-19 · Xinyao Zhang, Wenkai Dong, Yuxin Song, Bo Fang 외 arxiv

Current instruction-guided video editing models struggle to simultaneously balance precise semantic modifications with faithful motion preservation. While existing approaches rely on injecting explicit external priors (e…

Video Restoration

InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing

2025-10-09 · Haoran Yu, Yi Shi arxiv

Text-to-image diffusion models have shown great potential for image editing, with techniques such as text-based and object-dragging methods emerging as key approaches. However, each of these methods has inherent limitati…

Text-based Image EditingImage Reconstruction

MMM: Generative Masked Motion Model

2023-12-06 · CVPR 2024 1 · Ekkasit Pinyoanuntapong, Pu Wang, Minwoo Lee, Chen Chen

Recent advances in text-to-motion generation using diffusion and autoregressive models have shown promising results. However, these models often suffer from a trade-off between real-time performance, high fidelity, and m…

GPUmodelMotion Generationmotion in-betweening+1

DPE: Disentanglement of Pose and Expression for General Video Portrait Editing

2023-01-16 · CVPR 2023 1 · Youxin Pang, Yong Zhang, Weize Quan, Yanbo Fan 외

One-shot video-driven talking face generation aims at producing a synthetic talking video by transferring the facial motion from a video to an arbitrary portrait image. Head pose and facial expression are always entangle…

DisentanglementFace GenerationTalking Face GenerationVideo Editing