Unsupervised Style-based Explicit 3D Face Reconstruction from Single Image
Inferring 3D object structures from a single image is an ill-posed task due to depth ambiguity and occlusion. Typical resolutions in the literature include leveraging 2D or 3D ground truth for supervised learning, as well as imposing hand-crafted symmetry priors or using an implicit representation to hallucinate novel viewpoints for unsupervised methods. In this work, we propose a general adversarial learning framework for solving Unsupervised 2D to Explicit 3D Style Transfer (UE3DST). Specifically, we merge two architectures: the unsupervised explicit 3D reconstruction network of Wu et al.\ and the Generative Adversarial Network (GAN) named StarGAN-v2. We experiment across three facial datasets (Basel Face Model, 3DFAW and CelebA-HQ) and show that our solution is able to outperform well established solutions such as DepthNet in 3D reconstruction and Pix2NeRF in conditional style transfer, while we also justify the individual contributions of our model components via ablation. In contrast to the aforementioned baselines, our scheme produces features for explicit 3D rendering, which can be manipulated and utilized in downstream tasks.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Face Reconstruction3D ReconstructionFace ModelFace ReconstructionGenerative Adversarial NetworkStyle TransferSimilar Papers 제목 키워드 기반
Rethinking Content and Style: Exploring Bias for Unsupervised Disentanglement
Content and style (C-S) disentanglement intends to decompose the underlying explanatory factors of objects into two independent subspaces. From the unsupervised disentanglement perspective, we rethink content and style a…
3D ReconstructionDisentanglementImage ReconstructionInductive Bias+2Rich Syntactic and Semantic Information Helps Unsupervised Text Style Transfer
Text style transfer aims to change an input sentence to an output sentence by changing its text style while preserving the content. Previous efforts on unsupervised text style transfer only use the surface features of wo…
SentenceStyle TransferText Style TransferUnsupervised Text Style TransferFast 3D Reconstruction of Faces With Glasses
We present a method for the fast 3D face reconstruction of people wearing glasses. Our method explicitly and robustly models the case in which a face to be reconstructed is partially occluded by glasses. We propose a sim…
3D Face Reconstruction3D ReconstructionFace ReconstructionMExECON: Multi-view Extended Explicit Clothed humans Optimized via Normal integration
This work presents MExECON, a novel pipeline for 3D reconstruction of clothed human avatars from sparse multi-view RGB images. Building on the single-view method ECON, MExECON extends its capabilities to leverage multipl…
3D ReconstructionPose EstimationUnsupervised Style and Content Separation by Minimizing Mutual Information for Speech Synthesis
We present a method to generate speech from input text and a style vector that is extracted from a reference speech signal in an unsupervised manner, i.e., no style annotation, such as speaker information, is required. E…
DecoderSpeech Synthesis