paper-with-me

홈 › Papers

3D-Fixup: Advancing Photo Editing with 3D Priors

2025-05-15 · Yen-Chi Cheng, Krishna Kumar Singh, Jae Shin Yoon, Alex Schwing, LiangYan Gui, Matheus Gadelha, Paul Guerrero, Nanxuan Zhao

Despite significant advances in modeling image priors via diffusion models, 3D-aware image editing remains challenging, in part because the object is only specified via a single image. To tackle this challenge, we propose 3D-Fixup, a new framework for editing 2D images guided by learned 3D priors. The framework supports difficult editing situations such as object translation and 3D rotation. To achieve this, we leverage a training-based approach that harnesses the generative power of diffusion models. As video data naturally encodes real-world physical dynamics, we turn to video data for generating training data pairs, i.e., a source and a target frame. Rather than relying solely on a single trained model to infer transformations between source and target frames, we incorporate 3D guidance from an Image-to-3D model, which bridges this challenging task by explicitly projecting 2D information into 3D space. We design a data generation pipeline to ensure high-quality 3D guidance throughout training. Results show that by integrating these 3D priors, 3D-Fixup effectively supports complex, identity coherent 3D-aware edits, achieving high-quality results and advancing the application of diffusion models in realistic image manipulation. The code is provided at https://3dfixup.github.io/

📄 PDF Abstract BibTeX arXiv:2505.10566

Code (0)

등록된 구현이 없습니다.

Tasks

Image ManipulationImage to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos

2024-03-19 · Hadi AlZayer, Zhihao Xia, Xuaner Zhang, Eli Shechtman 외

We propose a generative model that, given a coarsely edited image, synthesizes a photorealistic output that follows the prescribed layout. Our method transfers fine details from the original image and preserves the ident…

Fixup Initialization: Residual Learning Without Normalization

2019-01-27 · ICLR 2019 5 · Hongyi Zhang, Yann N. Dauphin, Tengyu Ma

Normalization layers are a staple in state-of-the-art deep neural network architectures. They are widely believed to stabilize training, enable higher learning rate, accelerate convergence and improve generalization, tho…

General Classificationimage-classificationImage ClassificationMachine Translation+1

DiffusionRig: Learning Personalized Priors for Facial Appearance Editing

2023-04-13 · CVPR 2023 1 · Zheng Ding, Xuaner Zhang, Zhihao Xia, Lars Jebe 외

We address the problem of learning person-specific facial priors from a small number (e.g., 20) of portrait photos of the same person. This enables us to edit this specific person's facial appearance, such as expression …

Text-guided Image-and-Shape Editing and Generation: A Short Survey

2023-04-18 · Cheng-Kang Ted Chao, Yotam Gingold

Image and shape editing are ubiquitous among digital artworks. Graphics algorithms facilitate artists and designers to achieve desired editing intents without going through manually tedious retouching. In the recent adva…

Survey

ReX-Shot: Single-Image Rephotography via Geometry- and Camera-Grounded Generation

2026-08-19 · Ruiqi Zhang, Hao Zhu, Wenhao Zhang, Qi Zhang 외 arxiv

Single-image rephotography aims to synthesize new shots of a scene from a single reference image with specified viewpoints, focal lengths, and photographic effects, which are intrinsically coupled in imaging. Existing me…

3D Reconstruction