paper-with-me

Papers

GeoEdit: Geometry-Aware Object Editing via Dual-Branch Denoising

2026-06-29 · Yi He, Jiangming Wang, Xinyu Wang, Mark Fong, Songchun Zhang, Yuxuan Xue, Hai-Tao Zheng, Yue Ma arxiv

Precisely manipulating objects in a single photograph (translation, rotation, scaling) while obeying 3D physical constraints remains unsolved for diffusion-based editors. Current 2D methods lack spatial awareness and produce perspective violations. Forcing structural proxies into the latent space also disrupts variance homogeneity, and the resulting self-attention leakage leads to ghosting and background blur. The core difficulty is asymmetric: the relocated object must follow a rigid geometry, yet the uncovered background needs freedom to synthesize plausible content. We present GeoEdit, a training-free Lift-Manipulate-Render-Denoise pipeline that satisfies both constraints. We decouple scene and object in 3D, align them through point correspondence, and render a geometry-aligned proxy with a structural depth map. A Dual-Branch Denoising stage then refines this proxy: a video diffusion backbone preserves object identity, while 3D constraints are injected into the foreground within a narrow denoising window at matching noise variance (variance-homogeneous injection). The background denoises freely. Because the injected signal matches the native latent statistics, self-attention stays undisturbed. We also introduce GeoEditBench, a pose-aware benchmark covering object translation, object rotation, and camera movement with pose-aware evaluation metrics. Experiments confirm consistent gains in geometric accuracy, identity fidelity, and background quality. Our codes are available at https://github.com/Heey731/GeoEdit.

📄 PDF Abstract BibTeX arXiv:2606.30003

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GeoEdit: Geometric Knowledge Editing for Large Language Models

2025-02-27 · Yujie Feng, LiMing Zhan, Zexin Lu, Yongxin Xu 외

Regular updates are essential for maintaining up-to-date knowledge in large language models (LLMs). Consequently, various model editing methods have been developed to update specific knowledge within LLMs. However, train…

General Knowledgeknowledge editingModel Editing

Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers

2026-02-09 · Shuo Zhang, Wenzhuo Wu, Huayu Zhang, Jiarong Cheng 외 arxiv

Recent advances in diffusion models have significantly improved image editing. However, challenges persist in handling geometric transformations, such as translation, rotation, and scaling, particularly in complex scenes…

Image Editing

Diffusion Models are Geometry Critics: Single Image 3D Editing Using Pre-Trained Diffusion Priors

2024-03-18 · Ruicheng Wang, Jianfeng Xiang, Jiaolong Yang, Xin Tong

We propose a novel image editing technique that enables 3D manipulations on single images, such as object rotation and translation. Existing 3D-aware image editing approaches typically rely on synthetic multi-view datase…

Novel View Synthesis

OBJECT 3DIT: Language-guided 3D-aware Image Editing

2023-07-20 · NeurIPS 2023 11

Existing image editing tools, while powerful, typically disregard the underlying 3D geometry from which the image is projected. As a result, edits made using these tools may become detached from the geometry and lighting…

3D geometryObject

Designing a 3D-Aware StyleNeRF Encoder for Face Editing

2023-02-19 · Songlin Yang, Wei Wang, Bo Peng, Jing Dong

GAN inversion has been exploited in many face manipulation tasks, but 2D GANs often fail to generate multi-view 3D consistent images. The encoders designed for 2D GANs are not able to provide sufficient 3D information fo…

AttributeFace ModelVideo Editing