paper-with-me

홈 › Papers

BodyGAN: General-Purpose Controllable Neural Human Body Generation

2022-01-01 · CVPR 2022 1 · Chaojie Yang, Hanhui Li, Shengjie Wu, Shengkai Zhang, Haonan Yan, Nianhong Jiao, Jie Tang, Runnan Zhou, Xiaodan Liang, Tianxiang Zheng

Recent advances in generative adversarial networks (GANs) have provided potential solutions for photorealistic human image synthesis. However, the explicit and individual control of synthesis over multiple factors, such as poses, body shapes, and skin colors, remains difficult for existing methods. This is because current methods mainly rely on a single pose/appearance model, which is limited in disentangling various poses and appearance in human images. In addition, such a unimodal strategy is prone to causing severe artifacts in the generated images like color distortions and unrealistic textures. To tackle these issues, this paper proposes a multi-factor conditioned method dubbed BodyGAN. Specifically, given a source image, our Body-GAN aims at capturing the characteristics of the human body from multiple aspects: (i) A pose encoding branch consisting of three hybrid subnetworks is adopted, to generate the semantic segmentation based representation, the 3D surface based representation, and the key point based representation of the human body, respectively. (ii) Based on the segmentation results, an appearance encoding branch is used to obtain the appearance information of the human body parts. (iii) The outputs of these two branches are represented by user-editable condition maps, which are then processed by a generator to predict the synthesized image. In this way, our BodyGAN can achieve the fine-grained disentanglement of pose, body shape, and appearance, and consequently enable the explicit and effective control of synthesis with diverse conditions. Extensive experiments on multiple datasets and a comprehensive user-study show that our BodyGAN achieves the state-of-the-art performance.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DisentanglementImage GenerationSemantic Segmentation

Similar Papers 제목 키워드 기반

The Wanderings of Odysseus in 3D Scenes

2021-12-16 · CVPR 2022 1 · Yan Zhang, Siyu Tang

Our goal is to populate digital environments, in which digital humans have diverse body shapes, move perpetually, and have plausible body-scene contact. The core challenge is to generate realistic, controllable, and infi…

Learning Disentangled Representations for Controllable Human Motion Prediction

2022-07-04 · Chunzhi Gu, Jun Yu, Chao Zhang

Generative model-based motion prediction techniques have recently realized predicting controlled human motions, such as predicting multiple upper human body motions with similar lower-body motions. However, to achieve th…

Human motion predictionInductive Biasmotion predictionPrediction

Towards Human-level Intelligence via Human-like Whole-Body Manipulation

2025-07-23 · Guang Gao, Jianan Wang, Jinbo Zuo, Junnan Jiang 외 arxiv

Building general-purpose intelligent robots has long been a fundamental goal of robotics. A promising approach is to mirror the evolutionary trajectory of humans: learning through continuous interaction with the environm…

Narrator: Towards Natural Control of Human-Scene Interaction Generation via Relationship Reasoning

2023-03-16 · ICCV 2023 1 · Haibiao Xuan, Xiongzheng Li, Jinsong Zhang, Hongwen Zhang 외

Naturally controllable human-scene interaction (HSI) generation has an important role in various fields, such as VR/AR content creation and human-centered AI. However, existing methods are unnatural and unintuitive in th…

Disentangled Human Body Representation Based on Unsupervised Semantic-Aware Learning

2025-05-25 · Lu Wang, Xishuai Peng, S. Kevin Zhou

In recent years, more and more attention has been paid to the learning of 3D human representation. However, the complexity of lots of hand-defined human body constraints and the absence of supervision data limit that the…

DisentanglementPose Transfer