paper-with-me

홈 › Papers

Normalized Avatar Synthesis Using StyleGAN and Perceptual Refinement

2021-06-21 · CVPR 2021 1 · Huiwen Luo, Koki Nagano, Han-Wei Kung, Mclean Goldwhite, Qingguo Xu, Zejian Wang, Lingyu Wei, Liwen Hu, Hao Li

We introduce a highly robust GAN-based framework for digitizing a normalized 3D avatar of a person from a single unconstrained photo. While the input image can be of a smiling person or taken in extreme lighting conditions, our method can reliably produce a high-quality textured model of a person's face in neutral expression and skin textures under diffuse lighting condition. Cutting-edge 3D face reconstruction methods use non-linear morphable face models combined with GAN-based decoders to capture the likeness and details of a person but fail to produce neutral head models with unshaded albedo textures which is critical for creating relightable and animation-friendly avatars for integration in virtual environments. The key challenges for existing methods to work is the lack of training and ground truth data containing normalized 3D faces. We propose a two-stage approach to address this problem. First, we adopt a highly robust normalized 3D face generator by embedding a non-linear morphable face model into a StyleGAN2 network. This allows us to generate detailed but normalized facial assets. This inference is then followed by a perceptual refinement step that uses the generated assets as regularization to cope with the limited available training samples of normalized faces. We further introduce a Normalized Face Dataset, which consists of a combination photogrammetry scans, carefully selected photographs, and generated fake people with neutral expressions in diffuse lighting conditions. While our prepared dataset contains two orders of magnitude less subjects than cutting edge GAN-based 3D facial reconstruction methods, we show that it is possible to produce high-quality normalized face models for very challenging unconstrained input images, and demonstrate superior performance to the current state-of-the-art.

📄 PDF Abstract BibTeX arXiv:2106.11423

Code (0)

등록된 구현이 없습니다.

Tasks

3D Face ReconstructionFace ModelFace Reconstruction

Methods 이 논문이 사용한 방법론

HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Path Length Regularization 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Weight Demodulation 설명 없음

Similar Papers 제목 키워드 기반

ToonifyGB: StyleGAN-based Gaussian Blendshapes for 3D Stylized Head Avatars

2025-05-15 · Rui-Yang Ju, Sheng-Yen Huang, Yi-Ping Hung

The introduction of 3D Gaussian blendshapes has enabled the real-time reconstruction of animatable head avatars from monocular video. Toonify, a StyleGAN-based framework, has become widely used for facial image stylizati…

Image StylizationVideo Generation

Texture Representation via Analysis and Synthesis with Generative Adversarial Networks

2022-12-20 · Jue Lin, Gaurav Sharma, Thrasyvoulos N. Pappas

We investigate data-driven texture modeling via analysis and synthesis with generative adversarial networks. For network training and testing, we have compiled a diverse set of spatially homogeneous textures, ranging fro…

Texture Classification

GaussianStyle: Gaussian Head Avatar via StyleGAN

2024-02-01 · Pinxin Liu, Luchuan Song, Daoan Zhang, Hang Hua 외

Existing methods like Neural Radiation Fields (NeRF) and 3D Gaussian Splatting (3DGS) have made significant strides in facial attribute control such as facial animation and components editing, yet they struggle with fine…

3DGSAttributeContrastive LearningNeRF+2

Self-Distilled StyleGAN: Towards Generation from Internet Photos

2022-02-24 · Ron Mokady, Michal Yarom, Omer Tov, Oran Lang 외

StyleGAN is known to produce high-fidelity images, while also offering unprecedented semantic editing. However, these fascinating abilities have been demonstrated only on a limited set of datasets, which are usually stru…

Image Generation

VisualSpeaker: Visually-Guided 3D Avatar Lip Synthesis

2025-07-08 · Alexandre Symeonidis-Herzig, Özge Mercanoğlu Sincan, Richard Bowden

Realistic, high-fidelity 3D facial animations are crucial for expressive avatar systems in human-computer interaction and accessibility. Although prior methods show promising quality, their reliance on the mesh domain li…

Automatic Speech RecognitionLip Readingspeech-recognitionSpeech Recognition+1