paper-with-me

Papers

V2Edit: Versatile Video Diffusion Editor for Videos and 3D Scenes

2025-03-13 · YanMing Zhang, Jun-Kun Chen, Jipeng Lyu, Yu-Xiong Wang

This paper introduces V$^2$Edit, a novel training-free framework for instruction-guided video and 3D scene editing. Addressing the critical challenge of balancing original content preservation with editing task fulfillment, our approach employs a progressive strategy that decomposes complex editing tasks into a sequence of simpler subtasks. Each subtask is controlled through three key synergistic mechanisms: the initial noise, noise added at each denoising step, and cross-attention maps between text prompts and video content. This ensures robust preservation of original video elements while effectively applying the desired edits. Beyond its native video editing capability, we extend V$^2$Edit to 3D scene editing via a "render-edit-reconstruct" process, enabling high-quality, 3D-consistent edits even for tasks involving substantial geometric changes such as object insertion. Extensive experiments demonstrate that our V$^2$Edit achieves high-quality and successful edits across various challenging video editing tasks and complex 3D scene editing tasks, thereby establishing state-of-the-art performance in both domains.

📄 PDF Abstract BibTeX arXiv:2503.10634

Code (0)

등록된 구현이 없습니다.

Tasks

3D scene EditingDenoisingVideo Editing

Similar Papers 제목 키워드 기반

VRWKV-Editor: Reducing quadratic complexity in transformer-based video editing

2025-09-30 · Abdelilah Aitrouga, Youssef Hmamouche, Amal El Fallah Seghrouchni arxiv

In light of recent progress in video editing, deep learning models focusing on both spatial and temporal dependencies have emerged as the primary method. However, these models suffer from the quadratic computational comp…

RoomEditor++: A Parameter-Sharing Diffusion Architecture for High-Fidelity Furniture Synthesis

2025-12-19 · Qilong Wang, Xiaofan Ming, Zhenyi Lin, Jinwen Li 외 arxiv

Virtual furniture synthesis, which seamlessly integrates reference objects into indoor scenes while maintaining geometric coherence and visual realism, holds substantial promise for home design and e-commerce application…

VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation

2025-03-18 · Shoubin Yu, Difan Liu, Ziqiao Ma, Yicong Hong 외

Recent video diffusion models have enhanced video editing, but it remains challenging to handle instructional editing and diverse tasks (e.g., adding, removing, changing) within a unified framework. In this paper, we int…

Reasoning SegmentationVideo Editing

ExpressEdit: Video Editing with Natural Language and Sketching

2024-03-26 · Bekzat Tilekbay, Saelyne Yang, Michal Lewkowicz, Alex Suryapranata 외

Informational videos serve as a crucial source for explaining conceptual and procedural knowledge to novices and experts alike. When producing informational videos, editors edit videos by overlaying text/images or trimmi…

Video Editing

Dreamix: Video Diffusion Models are General Video Editors

2023-02-02 · Eyal Molad, Eliahu Horwitz, Dani Valevski, Alex Rav Acha 외

Text-driven image and video diffusion models have recently achieved unprecedented generation realism. While diffusion models have been successfully applied for image editing, very few works have done so for video editing…

Image AnimationImage to Video GenerationSubject-driven Video GenerationText-to-Video Editing+2