paper-with-me

Papers

LatentSwap3D: Semantic Edits on 3D Image GANs

2022-12-02 · Enis Simsar, Alessio Tonioni, Evin Pınar Örnek, Federico Tombari

3D GANs have the ability to generate latent codes for entire 3D volumes rather than only 2D images. These models offer desirable features like high-quality geometry and multi-view consistency, but, unlike their 2D counterparts, complex semantic image editing tasks for 3D GANs have only been partially explored. To address this problem, we propose LatentSwap3D, a semantic edit approach based on latent space discovery that can be used with any off-the-shelf 3D or 2D GAN model and on any dataset. LatentSwap3D relies on identifying the latent code dimensions corresponding to specific attributes by feature ranking using a random forest classifier. It then performs the edit by swapping the selected dimensions of the image being edited with the ones from an automatically selected reference image. Compared to other latent space control-based edit methods, which were mainly designed for 2D GANs, our method on 3D GANs provides remarkably consistent semantic edits in a disentangled manner and outperforms others both qualitatively and quantitatively. We show results on seven 3D GANs (pi-GAN, GIRAFFE, StyleSDF, MVCGAN, EG3D, StyleNeRF, and VolumeGAN) and on five datasets (FFHQ, AFHQ, Cats, MetFaces, and CompCars).

📄 PDF Abstract BibTeX arXiv:2212.01381

Code (0)

등록된 구현이 없습니다.

Tasks

Feature Importance

Methods 이 논문이 사용한 방법론

HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
StyleGAN 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Adaptive Instance Normalization 설명 없음
Feedforward Network A Feedforward Network, or a Multilayer Perceptron (MLP), is a neural network with solely densely connected layers. This is the classic neural network architecture of the…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

$S^2$-Flow: Joint Semantic and Style Editing of Facial Images

2022-11-22 · Krishnakant Singh, Simone Schaub-Meyer, Stefan Roth

The high-quality images yielded by generative adversarial networks (GANs) have motivated investigations into their application for image editing. However, GANs are often limited in the control they provide for performing…

Decoder

LatentSwap: An Efficient Latent Code Mapping Framework for Face Swapping

2024-02-28 · Changho Choi, Minho Kim, Junhyeok Lee, Hyoung-Kyu Song 외

We propose LatentSwap, a simple face swapping framework generating a face swap latent code of a given generator. Utilizing randomly sampled latent codes, our framework is light and does not require datasets besides emplo…

Face Swapping

Editing in Style: Uncovering the Local Semantics of GANs

2020-04-29 · CVPR 2020 6 · Edo Collins, Raja Bala, Bob Price, Sabine Süsstrunk

While the quality of GAN image synthesis has improved tremendously in recent years, our ability to control and condition the output is still limited. Focusing on StyleGAN, we introduce a simple and effective method for m…

DisentanglementImage Generation

Wasserstein Loss for Semantic Editing in the Latent Space of GANs

2023-03-22 · Perla Doubinsky, Nicolas Audebert, Michel Crucianu, Hervé Le Borgne

The latent space of GANs contains rich semantics reflecting the training data. Different methods propose to learn edits in latent space corresponding to semantic attributes, thus allowing to modify generated images. Most…

VIVE3D: Viewpoint-Independent Video Editing using 3D-Aware GANs

2023-03-28 · CVPR 2023 1 · Anna Frühstück, Nikolaos Sarafianos, Yuanlu Xu, Peter Wonka 외

We introduce VIVE3D, a novel approach that extends the capabilities of image-based 3D GANs to video editing and is able to represent the input video in an identity-preserving and temporally consistent way. We propose two…

Optical Flow EstimationVideo Editing