Attribute2Image: Conditional Image Generation from Visual Attributes
This paper investigates a novel problem of generating images from visual attributes. We model the image as a composite of foreground and background and develop a layered generative model with disentangled latent variables that can be learned end-to-end using a variational auto-encoder. We experiment with natural images of faces and birds and demonstrate that the proposed models are capable of generating realistic and diverse samples with disentangled latent representations. We use a general energy minimization algorithm for posterior inference of latent variables given novel images. Therefore, the learned generative models show excellent quantitative and visual results in the tasks of attribute-conditioned image reconstruction and completion.
Code (1)
Tasks
AttributeConditional Image GenerationImage GenerationImage ReconstructionSimilar Papers 제목 키워드 기반
Multi-attribute Pizza Generator: Cross-domain Attribute Control with Conditional StyleGAN
Multi-attribute conditional image generation is a challenging problem in computervision. We propose Multi-attribute Pizza Generator (MPG), a conditional Generative Neural Network (GAN) framework for synthesizing images f…
AttributeConditional Image GenerationImage GenerationAttribute-Guided Face Generation Using Conditional CycleGAN
We are interested in attribute-guided face generation: given a low-res face input image, an attribute vector that can be extracted from a high-res image (attribute image), our new method generates a high-res face image f…
AttributeFace GenerationFace SwappingFace VerificationMichiGAN: Multi-Input-Conditioned Hair Image Generation for Portrait Editing
Despite the recent success of face image generation with GANs, conditional hair editing remains challenging due to the under-explored complexity of its geometry and appearance. In this paper, we present MichiGAN (Multi-I…
Conditional Image GenerationImage GenerationArtVLM: Attribute Recognition Through Vision-Based Prefix Language Modeling
Recognizing and disentangling visual attributes from objects is a foundation to many computer vision applications. While large vision language representations like CLIP had largely resolved the task of zero-shot object r…
AttributeLanguage ModelingLanguage ModellingObject+4Localizing and Editing Knowledge in Text-to-Image Generative Models
Text-to-Image Diffusion Models such as Stable-Diffusion and Imagen have achieved unprecedented quality of photorealism with state-of-the-art FID scores on MS-COCO and other generation benchmarks. Given a caption, image g…
AttributeImage GenerationModel Editing