FineGAN: Unsupervised Hierarchical Disentanglement for Fine-Grained Object Generation and Discovery
We propose FineGAN, a novel unsupervised GAN framework, which disentangles the background, object shape, and object appearance to hierarchically generate images of fine-grained object categories. To disentangle the factors without supervision, our key idea is to use information theory to associate each factor to a latent code, and to condition the relationships between the codes in a specific way to induce the desired hierarchy. Through extensive experiments, we show that FineGAN achieves the desired disentanglement to generate realistic and diverse images belonging to fine-grained classes of birds, dogs, and cars. Using FineGAN's automatically learned features, we also cluster real images as a first attempt at solving the novel problem of unsupervised fine-grained object category discovery. Our code/models/demo can be found at https://github.com/kkanshul/finegan
Code (1)
Tasks
Conditional Image GenerationDisentanglementFine-Grained Visual CategorizationImage ClusteringObjectMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
cFineGAN: Unsupervised multi-conditional fine-grained image generation
We propose an unsupervised multi-conditional image generation pipeline: cFineGAN, that can generate an image conditioned on two input images such that the generated image preserves the texture of one and the shape of the…
Conditional Image GenerationImage GenerationMixNMatch: Multifactor Disentanglement and Encoding for Conditional Image Generation
We present MixNMatch, a conditional generative model that learns to disentangle and encode background, object pose, shape, and texture from real images with minimal supervision, for mix-and-match image generation. We bui…
Conditional Image GenerationDisentanglementImage GenerationClinically Plausible Pathology-Anatomy Disentanglement in Patient Brain MRI with Structured Variational Priors
We propose a hierarchically structured variational inference model for accurately disentangling observable evidence of disease (e.g. brain lesions or atrophy) from subject-specific anatomy in brain MRIs. With flexible, p…
AnatomyDisentanglementVariational InferenceRetrieve in Style: Unsupervised Facial Feature Transfer and Retrieval
We present Retrieve in Style (RIS), an unsupervised framework for facial feature transfer and retrieval on real images. Recent work shows capabilities of transferring local facial features by capitalizing on the disentan…
DisentanglementRetrievalRefineGAN: Universally Generating Waveform Better than Ground Truth with Highly Accurate Pitch and Intensity Responses
Most GAN(Generative Adversarial Network)-based approaches towards high-fidelity waveform generation heavily rely on discriminators to improve their performance. However, GAN methods introduce much uncertainty into the ge…
Audio GenerationGenerative Adversarial NetworkSinging Voice Synthesis