paper-with-me

홈 › Papers

Omni-3DEdit: Generalized Versatile 3D Editing in One-Pass

2026-03-18 · Chen Liyi, Wang Pengfei, Zhang Guowen, Ma Zhiyuan, Zhang Lei arxiv

Most instruction-driven 3D editing methods rely on 2D models to guide the explicit and iterative optimization of 3D representations. This paradigm, however, suffers from two primary drawbacks. First, it lacks a universal design of different 3D editing tasks because the explicit manipulation of 3D geometry necessitates task-dependent rules, e.g., 3D appearance editing demands inherent source 3D geometry, while 3D removal alters source geometry. Second, the iterative optimization process is highly time-consuming, often requiring thousands of invocations of 2D/3D updating. We present Omni-3DEdit, a unified, learning-based model that generalizes various 3D editing tasks implicitly. One key challenge to achieve our goal is the scarcity of paired source-edited multi-view assets for training. To address this issue, we construct a data pipeline, synthesizing a relatively rich number of high-quality paired multi-view editing samples. Subsequently, we adapt the pre-trained generative model SEVA as our backbone by concatenating source view latents along with conditional tokens in sequence space. A dual-stream LoRA module is proposed to disentangle different view cues, largely enhancing our model's representational learning capability. As a learning-based model, our model is free of the time-consuming online optimization, and it can complete various 3D editing tasks in one forward pass, reducing the inference time from tens of minutes to approximately two minutes. Extensive experiments demonstrate the effectiveness and efficiency of Omni-3DEdit.

📄 PDF Abstract BibTeX arXiv:2603.17841

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

OmniGuard: Hybrid Manipulation Localization via Augmented Versatile Deep Image Watermarking

2024-12-02 · CVPR 2025 1 · Xuanyu Zhang, Zecheng Tang, Zhipei Xu, Runyi Li 외

With the rapid growth of generative AI and its widespread application in image editing, new risks have emerged regarding the authenticity and integrity of digital content. Existing versatile watermarking approaches suffe…

MedEdit: Counterfactual Diffusion-based Image Editing on Brain MRI

2024-07-21 · Malek Ben Alaya, Daniel M. Lang, Benedikt Wiestler, Julia A. Schnabel 외

Denoising diffusion probabilistic models enable high-fidelity image synthesis and editing. In biomedicine, these models facilitate counterfactual image editing, producing pairs of images where one is edited to simulate h…

counterfactualDenoisingImage Generation

OmniGen2: Exploration to Advanced Multimodal Generation

2025-06-23 · Chenyuan Wu, Pengfei Zheng, Ruiran Yan, Shitao Xiao 외

In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generation tasks, including text-to-image, image editing, and in-context generation. Un…

Image Generationmultimodal generationText Generation

Rethinking One-Step Image Editing through ChordEdit: Reproduction, Simplification, and New Insights

2026-06-12 · Minghan Li, Jeremy Moebel, Mengyu Wang arxiv

One-step image editing is important for making text-guided editing fast, practical, and easy to deploy, but its underlying mechanism is still not fully understood. We revisit ChordEdit through reproduction, ablation, and…

Image Editing

BadEdit: Backdooring large language models by model editing

2024-03-20 · Yanzhou Li, Tianlin Li, Kangjie Chen, Jian Zhang 외

Mainstream backdoor attack methods typically demand substantial tuning data for poisoning, limiting their practicality and potentially degrading the overall performance when applied to Large Language Models (LLMs). To ad…

Backdoor Attackknowledge editingmodelModel Editing