paper-with-me

Papers

NeRFEditor: Differentiable Style Decomposition for Full 3D Scene Editing

2022-12-07 · Chunyi Sun, Yanbin Liu, Junlin Han, Stephen Gould

We present NeRFEditor, an efficient learning framework for 3D scene editing, which takes a video captured over 360{\deg} as input and outputs a high-quality, identity-preserving stylized 3D scene. Our method supports diverse types of editing such as guided by reference images, text prompts, and user interactions. We achieve this by encouraging a pre-trained StyleGAN model and a NeRF model to learn from each other mutually. Specifically, we use a NeRF model to generate numerous image-angle pairs to train an adjustor, which can adjust the StyleGAN latent code to generate high-fidelity stylized images for any given angle. To extrapolate editing to GAN out-of-domain views, we devise another module that is trained in a self-supervised learning manner. This module maps novel-view images to the hidden space of StyleGAN that allows StyleGAN to generate stylized images on novel views. These two modules together produce guided images in 360{\deg}views to finetune a NeRF to make stylization effects, where a stable fine-tuning strategy is proposed to achieve this. Experiments show that NeRFEditor outperforms prior work on benchmark and real-world scenes with better editability, fidelity, and identity preservation.

📄 PDF Abstract BibTeX arXiv:2212.03848

Code (0)

등록된 구현이 없습니다.

Tasks

3D scene EditingNeRFSelf-Supervised Learning

Methods 이 논문이 사용한 방법론

StyleGAN 설명 없음
Adaptive Instance Normalization 설명 없음
HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Feedforward Network A Feedforward Network, or a Multilayer Perceptron (MLP), is a neural network with solely densely connected layers. This is the classic neural network architecture of the…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

NeAI: A Pre-convoluted Representation for Plug-and-Play Neural Ambient Illumination

2023-04-18 · Yiyu Zhuang, Qi Zhang, Xuan Wang, Hao Zhu 외

Recent advances in implicit neural representation have demonstrated the ability to recover detailed geometry and material from multi-view images. However, the use of simplified lighting models such as environment maps to…

NeRF

Differentiable Blocks World: Qualitative 3D Decomposition by Rendering Primitives

2023-07-11 · NeurIPS 2023 11 · Tom Monnier, Jake Austin, Angjoo Kanazawa, Alexei A. Efros 외

Given a set of calibrated images of a scene, we present an approach that produces a simple, compact, and actionable 3D world representation by means of 3D primitives. While many approaches focus on recovering high-fideli…

Physical Simulations

RewriteNet: Reliable Scene Text Editing with Implicit Decomposition of Text Contents and Styles

2021-07-23 · Junyeop Lee, Yoonsik Kim, Seonghyeon Kim, Moonbin Yim 외

Scene text editing (STE), which converts a text in a scene image into the desired text while preserving an original style, is a challenging task due to a complex intervention between text and style. In this paper, we pro…

Image GenerationScene Text EditingScene Text Recognition

SILT: Self-supervised Lighting Transfer Using Implicit Image Decomposition

2021-10-25 · Nikolina Kubiak, Armin Mustafa, Graeme Phillipson, Stephen Jolly 외

We present SILT, a Self-supervised Implicit Lighting Transfer method. Unlike previous research on scene relighting, we do not seek to apply arbitrary new lighting configurations to a given scene. Instead, we wish to tran…

Towards Full-to-Empty Room Generation with Structure-Aware Feature Encoding and Soft Semantic Region-Adaptive Normalization

2021-12-10 · Vasileios Gkitsas, Nikolaos Zioulis, Vladimiros Sterzentsenko, Alexandros Doumanoglou 외

The task of transforming a furnished room image into a background-only is extremely challenging since it requires making large changes regarding the scene context while still preserving the overall layout and style. In o…

Depth EstimationImage InpaintingLayout Generation