paper-with-me

홈 › Papers

Neural Video Fields Editing

2023-12-12 · Shuzhou Yang, Chong Mou, Jiwen Yu, YuHan Wang, Xiandong Meng, Jian Zhang

Diffusion models have revolutionized text-driven video editing. However, applying these methods to real-world editing encounters two significant challenges: (1) the rapid increase in GPU memory demand as the number of frames grows, and (2) the inter-frame inconsistency in edited videos. To this end, we propose NVEdit, a novel text-driven video editing framework designed to mitigate memory overhead and improve consistent editing for real-world long videos. Specifically, we construct a neural video field, powered by tri-plane and sparse grid, to enable encoding long videos with hundreds of frames in a memory-efficient manner. Next, we update the video field through off-the-shelf Text-to-Image (T2I) models to impart text-driven editing effects. A progressive optimization strategy is developed to preserve original temporal priors. Importantly, both the neural video field and T2I model are adaptable and replaceable, thus inspiring future research. Experiments demonstrate the ability of our approach to edit hundreds of frames with impressive inter-frame consistency. Our project is available at: https://nvedit.github.io/.

📄 PDF Abstract BibTeX arXiv:2312.08882

Code (0)

등록된 구현이 없습니다.

Tasks

GPUVideo Editing

Similar Papers 제목 키워드 기반

DynVideo-E: Harnessing Dynamic NeRF for Large-Scale Motion- and View-Change Human-Centric Video Editing

2023-10-16 · CVPR 2024 1 · Jia-Wei Liu, Yan-Pei Cao, Jay Zhangjie Wu, Weijia Mao 외

Despite recent progress in diffusion-based video editing, existing methods are limited to short-length videos due to the contradiction between long-range consistency and frame-wise editing. Prior attempts to address this…

NeRFStyle TransferSuper-ResolutionVideo Editing

Pathways on the Image Manifold: Image Editing via Video Generation

2024-11-25 · CVPR 2025 1 · Noam Rotstein, Gal Yona, Daniel Silver, Roy Velich 외

Recent advances in image editing, driven by image diffusion models, have shown remarkable progress. However, significant challenges remain, as these models often struggle to follow complex edit instructions accurately an…

Text-based Image EditingVideo Generation

VideoSPatS: Video SPatiotemporal Splines for Disentangled Occlusion, Appearance and Motion Modeling and Editing

2025-04-08 · CVPR 2025 1 · Juan Luis Gonzalez Bello, Xu Yao, Alex Whelan, Kyle Olszewski 외

We present an implicit video representation for occlusions, appearance, and motion disentanglement from monocular videos, which we call Video SPatiotemporal Splines (VideoSPatS). Unlike previous methods that map time and…

DisentanglementMotion DisentanglementVideo Editing

OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation

2025-06-02 · Sen Liang, Zhentao Yu, Zhengguang Zhou, Teng Hu 외

The emergence of Diffusion Transformers (DiT) has brought significant advancements to video generation, especially in text-to-video and image-to-video tasks. Although video generation is widely applied in various fields,…

Data AugmentationHuman AnimationVideo Generation

MLV-Edit: Towards Consistent and Highly Efficient Editing for Minute-Level Videos

2026-02-02 · Yangyi Cao, Yuanhang Li, Lan Chen, Qi Mao arxiv

We propose MLV-Edit, a training-free, flow-based framework that address the unique challenges of minute-level video editing. While existing techniques excel in short-form video manipulation, scaling them to long-duration…