paper-with-me

Papers

SHAP-EDITOR: Instruction-guided Latent 3D Editing in Seconds

2023-12-14 · CVPR 2024 1 · Minghao Chen, Junyu Xie, Iro Laina, Andrea Vedaldi

We propose a novel feed-forward 3D editing framework called Shap-Editor. Prior research on editing 3D objects primarily concentrated on editing individual objects by leveraging off-the-shelf 2D image editing networks. This is achieved via a process called distillation, which transfers knowledge from the 2D network to 3D assets. Distillation necessitates at least tens of minutes per asset to attain satisfactory editing results, and is thus not very practical. In contrast, we ask whether 3D editing can be carried out directly by a feed-forward network, eschewing test-time optimisation. In particular, we hypothesise that editing can be greatly simplified by first encoding 3D objects in a suitable latent space. We validate this hypothesis by building upon the latent space of Shap-E. We demonstrate that direct 3D editing in this space is possible and efficient by building a feed-forward editor network that only requires approximately one second per edit. Our experiments show that Shap-Editor generalises well to both in-distribution and out-of-distribution 3D assets with different prompts, exhibiting comparable performance with methods that carry out test-time optimisation for each edited instance.

📄 PDF Abstract BibTeX arXiv:2312.09246

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ShapeWalk: Compositional Shape Editing Through Language-Guided Chains

2024-01-01 · CVPR 2024 1 · Habib Slim, Mohamed Elhoseiny

Editing 3D shapes through natural language instructions is a challenging task that requires the comprehension of both language semantics and fine-grained geometric details. To bridge this gap we introduce ShapeWalk a…

One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing

2026-09-03 · Adheesh Sunil Juvekar, Onkar Kishor Susladkar, Kiet A. Nguyen, Muntasir Wahed 외 hf

Video editing spans diverse editing paradigms, yet achieving high-quality instruction-guided and subject-guided editing within a single unified framework remains challenging. We introduce EditVid, a training-free framewo…

Style Transfer

Instruction-Based Video Editing by Repurposing an Image Editing Model

2026-08-14 · Yunpeng Bai, Yossi Gandelsman, Michaël Gharbi, Qixing Huang arxiv

Instruction-based video editing is commonly built on video-pretrained generative backbones: a video diffusion transformer is adapted, at considerable cost, to condition on a source video and an editing instruction. In th…

Image Editing

Are Image-to-Video Models Good Zero-Shot Image Editors?

2025-11-24 · Zechuan Zhang, Zhenyuan Chen, Zongxin Yang, Yi Yang arxiv

Large-scale video diffusion models show strong world simulation and temporal reasoning abilities, but their use as zero-shot image editors remains underexplored. We introduce IF-Edit, a tuning-free framework that repurpo…

Image Editing

NeuralEditor: Editing Neural Radiance Fields via Manipulating Point Clouds

2023-05-04 · CVPR 2023 1 · Jun-Kun Chen, Jipeng Lyu, Yu-Xiong Wang

This paper proposes NeuralEditor that enables neural radiance fields (NeRFs) natively editable for general shape editing tasks. Despite their impressive results on novel-view synthesis, it remains a fundamental challenge…

NeRFNovel View Synthesis