paper-with-me

Papers

Disentangled Representation Learning for Controllable Person Image Generation

2023-12-10 · Wenju Xu, Chengjiang Long, Yongwei Nie, Guanghui Wang

In this paper, we propose a novel framework named DRL-CPG to learn disentangled latent representation for controllable person image generation, which can produce realistic person images with desired poses and human attributes (e.g., pose, head, upper clothes, and pants) provided by various source persons. Unlike the existing works leveraging the semantic masks to obtain the representation of each component, we propose to generate disentangled latent code via a novel attribute encoder with transformers trained in a manner of curriculum learning from a relatively easy step to a gradually hard one. A random component mask-agnostic strategy is introduced to randomly remove component masks from the person segmentation masks, which aims at increasing the difficulty of training and promoting the transformer encoder to recognize the underlying boundaries between each component. This enables the model to transfer both the shape and texture of the components. Furthermore, we propose a novel attribute decoder network to integrate multi-level attributes (e.g., the structure feature and the attribute representation) with well-designed Dual Adaptive Denormalization (DAD) residual blocks. Extensive experiments strongly demonstrate that the proposed approach is able to transfer both the texture and shape of different human parts and yield realistic results. To our knowledge, we are the first to learn disentangled latent representations with transformers for person image generation.

📄 PDF Abstract BibTeX arXiv:2312.05798

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeDecoderImage GenerationRepresentation Learning

Similar Papers 제목 키워드 기반

DRC: Enhancing Personalized Image Generation via Disentangled Representation Composition

2025-04-24 · Yiyan Xu, Wuqiang Zheng, Wenjie Wang, Fengbin Zhu 외

Personalized image generation has emerged as a promising direction in multimodal content creation. It aims to synthesize images tailored to individual style preferences (e.g., color schemes, character appearances, layout…

DisentanglementImage GenerationPersonalized Image GenerationRepresentation Learning

DRDM: A Disentangled Representations Diffusion Model for Synthesizing Realistic Person Images

2024-12-25 · Enbo Huang, Yuan Zhang, Faliang Huang, Guangyu Zhang 외

Person image synthesis with controllable body poses and appearances is an essential task owing to the practical needs in the context of virtual try-on, image editing and video production. However, existing methods face s…

Image GenerationPose TransferVirtual Try-on

Content-style disentangled representation for controllable artistic image stylization and generation

2024-12-19 · Ma Zhuoqi, Zhang Yixuan, You Zejun, Tian Long 외

Controllable artistic image stylization and generation aims to render the content provided by text or image with the learned artistic style, where content and style decoupling is the key to achieve satisfactory results. …

DisentanglementImage Stylization

Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive Learning

2020-04-24 · CVPR 2020 6 · Yu Deng, Jiaolong Yang, Dong Chen, Fang Wen 외

We propose DiscoFaceGAN, an approach for face image generation of virtual people with disentangled, precisely-controllable latent representations for identity of non-existing people, expression, pose, and illumination. W…

Contrastive LearningDisentanglementImage Generation

Semi-Supervised StyleGAN for Disentanglement Learning

2020-03-06 · ICML 2020 1 · Weili Nie, Tero Karras, Animesh Garg, Shoubhik Debnath 외

Disentanglement learning is crucial for obtaining disentangled representations and controllable generation. Current disentanglement methods face several inherent limitations: difficulty with high-resolution images, prima…

DisentanglementRepresentation Learning