paper-with-me

홈 › Papers

Generating Person Images with Appearance-aware Pose Stylizer

2020-07-17 · Siyu Huang, Haoyi Xiong, Zhi-Qi Cheng, Qingzhong Wang, Xingran Zhou, Bihan Wen, Jun Huan, Dejing Dou

Generation of high-quality person images is challenging, due to the sophisticated entanglements among image factors, e.g., appearance, pose, foreground, background, local details, global structures, etc. In this paper, we present a novel end-to-end framework to generate realistic person images based on given person poses and appearances. The core of our framework is a novel generator called Appearance-aware Pose Stylizer (APS) which generates human images by coupling the target pose with the conditioned person appearance progressively. The framework is highly flexible and controllable by effectively decoupling various complex person image factors in the encoding phase, followed by re-coupling them in the decoding phase. In addition, we present a new normalization method named adaptive patch normalization, which enables region-specific normalization and shows a good performance when adopted in person image generation model. Experiments on two benchmark datasets show that our method is capable of generating visually appealing and realistic-looking results using arbitrary image and pose inputs.

📄 PDF Abstract BibTeX arXiv:2007.09077

Code (1)

siyuhuang/PoseStylizer 공식 구현 pytorch

Tasks

Image Generation

Similar Papers 제목 키워드 기반

LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation

2024-11-22 · Fan Deng, Yaguang Wu, Xinyang Yu, Xiangjun Huang 외

Recently, text-to-image models based on diffusion have achieved remarkable success in generating high-quality images. However, the challenge of personalized, controllable generation of instances within these images remai…

SA-Person: Text-Based Person Retrieval with Scene-aware Re-ranking

2025-05-30 · Yingjia Xu, Jinlin Wu, Zhen Chen, Daming Gao 외

Text-based person retrieval aims to identify a target individual from a gallery of images based on a natural language description. It presents a significant challenge due to the complexity of real-world scenes and the am…

Cross-Modal RetrievalPerson RetrievalRe-RankingRetrieval+2

Visual Persona: Foundation Model for Full-Body Human Customization

2025-03-19 · CVPR 2025 1 · Jisu Nam, Soowon Son, Zhan Xu, Jing Shi 외

We introduce Visual Persona, a foundation model for text-to-image full-body human customization that, given a single in-the-wild human image, generates diverse images of the individual guided by text descriptions. Unlike…

Appearance Transfer

MetaDance: Few-shot Dancing Video Retargeting via Temporal-aware Meta-learning

2022-01-13 · Yuying Ge, Yibing Song, Ruimao Zhang, Ping Luo

Dancing video retargeting aims to synthesize a video that transfers the dance movements from a source video to a target person. Previous work need collect a several-minute-long video of a target person with thousands of …

Meta-Learning

PEGAsus: 3D Personalization of Geometry and Appearance

2026-02-09 · Jingyu Hu, Bin Hu, Ka-Hei Hui, Haipeng Li 외 arxiv

We present PEGAsus, a new framework capable of generating Personalized 3D shapes by learning shape concepts at both Geometry and Appearance levels. First, we formulate 3D shape personalization as extracting reusable, cat…