paper-with-me

Papers

Age-Oriented Face Synthesis with Conditional Discriminator Pool and Adversarial Triplet Loss

2020-07-01 · Haoyi Wang, Victor Sanchez, Chang-Tsun Li

The vanilla Generative Adversarial Networks (GAN) are commonly used to generate realistic images depicting aged and rejuvenated faces. However, the performance of such vanilla GANs in the age-oriented face synthesis task is often compromised by the mode collapse issue, which may result in the generation of faces with minimal variations and a poor synthesis accuracy. In addition, recent age-oriented face synthesis methods use the L1 or L2 constraint to preserve the identity information on synthesized faces, which implicitly limits the identity permanence capabilities when these constraints are associated with a trivial weighting factor. In this paper, we propose a method for the age-oriented face synthesis task that achieves a high synthesis accuracy with strong identity permanence capabilities. Specifically, to achieve a high synthesis accuracy, our method tackles the mode collapse issue with a novel Conditional Discriminator Pool (CDP), which consists of multiple discriminators, each targeting one particular age category. To achieve strong identity permanence capabilities, our method uses a novel Adversarial Triplet loss. This loss, which is based on the Triplet loss, adds a ranking operation to further pull the positive embedding towards the anchor embedding resulting in significantly reduced intra-class variances in the feature space. Through extensive experiments, we show that our proposed method outperforms state-of-the-art methods in terms of synthesis accuracy and identity permanence capabilities, qualitatively and quantitatively.

📄 PDF Abstract BibTeX arXiv:2007.00792

Code (0)

등록된 구현이 없습니다.

Tasks

Face GenerationTriplet

Similar Papers 제목 키워드 기반

Parallel waveform synthesis based on generative adversarial networks with voicing-aware conditional discriminators

2020-10-27 · Ryuichi Yamamoto, Eunwoo Song, Min-Jae Hwang, Jae-Min Kim

This paper proposes voicing-aware conditional discriminators for Parallel WaveGAN-based waveform synthesis systems. In this framework, we adopt a projection-based conditioning method that can significantly improve the di…

text-to-speechText to Speech

GAN You Hear Me? Reclaiming Unconditional Speech Synthesis from Diffusion Models

2022-10-11 · Matthew Baas, Herman Kamper

We propose AudioStyleGAN (ASGAN), a new generative adversarial network (GAN) for unconditional speech synthesis. As in the StyleGAN family of image synthesis models, ASGAN maps sampled noise to a disentangled latent vect…

DisentanglementGenerative Adversarial NetworkSpeech SynthesisVoice Conversion

Mask-Embedded Discriminator With Region-Based Semantic Regularization for Semi-Supervised Class-Conditional Image Synthesis

2021-06-19 · CVPR 2021 1 · Yi Liu, Xiaoyang Huo, Tianyi Chen, Xiangping Zeng 외

Semi-supervised generative learning (SSGL) makes use of unlabeled data to achieve a trade-off between the data collection/annotation effort and generation performance, when adequate labeled data are not available. Le…

Generative Adversarial NetworkImage Generation

NitroFusion: High-Fidelity Single-Step Diffusion through Dynamic Adversarial Training

2024-12-02 · CVPR 2025 1 · Dar-Yen Chen, Hmrishav Bandyopadhyay, Kai Zou, Yi-Zhe Song

We introduce NitroFusion, a fundamentally different approach to single-step diffusion that achieves high-quality generation through a dynamic adversarial framework. While one-step methods offer dramatic speed advantages,…

Denoising

Class-Continuous Conditional Generative Neural Radiance Field

2023-01-03 · Jiwook Kim, Minhyeok Lee

The 3D-aware image synthesis focuses on conserving spatial consistency besides generating high-resolution images with fine details. Recently, Neural Radiance Field (NeRF) has been introduced for synthesizing novel views …

3D-Aware Image SynthesisImage GenerationImage-to-Image TranslationNeRF