paper-with-me

홈 › Papers

Video synthesis of human upper body with realistic face

2019-08-19 · Zhaoxiang Liu, Huan Hu, Zipeng Wang, Kai Wang, Jinqiang Bai, Shiguo Lian

This paper presents a generative adversarial learning-based human upper body video synthesis approach to generate an upper body video of target person that is consistent with the body motion, face expression, and pose of the person in source video. We use upper body keypoints, facial action units and poses as intermediate representations between source video and target video. Instead of directly transferring the source video to the target video, we firstly map the source person's facial action units and poses into the target person's facial landmarks, then combine the normalized upper body keypoints and generated facial landmarks with spatio-temporal smoothing to generate the corresponding target video's image. Experimental results demonstrated the effectiveness of our method.

📄 PDF Abstract BibTeX arXiv:1908.06607

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers

2025-03-13 · Yasheng Sun, Zhiliang Xu, Hang Zhou, Jiazhi Guan 외

Co-speech gesture video synthesis is a challenging task that requires both probabilistic modeling of human gestures and the synthesis of realistic images that align with the rhythmic nuances of speech. To address these c…

Stereo-Talker: Audio-driven 3D Human Synthesis with Prior-Guided Mixture-of-Experts

2024-10-31 · Xiang Deng, Youxin Pang, Xiaochen Zhao, Chao Xu 외

This paper introduces Stereo-Talker, a novel one-shot audio-driven human video synthesis system that generates 3D talking videos with precise lip synchronization, expressive body gestures, temporally consistent photo-rea…

Language ModelingLanguage ModellingLarge Language ModelMixture-of-Experts+1

Physically Plausible Animation of Human Upper Body from a Single Image

2022-12-09 · Ziyuan Huang, Zhengping Zhou, Yung-Yu Chuang, Jiajun Wu 외

We present a new method for generating controllable, dynamically responsive, and photorealistic human animations. Given an image of a person, our system allows the user to generate Physically plausible Upper Body Animati…

Speech2Video Synthesis with 3D Skeleton Regularization and Expressive Body Poses

2020-07-17 · Miao Liao, Sibo Zhang, Peng Wang, Hao Zhu 외

In this paper, we propose a novel approach to convert given speech audio to a photo-realistic speaking video of a specific person, where the output video has synchronized, realistic, and expressive rich body dynamics. We…

Generative Adversarial Network

Towards More Realistic Human-Robot Conversation: A Seq2Seq-based Body Gesture Interaction System

2019-05-05 · Minjie Hua, Fuyuan Shi, Yibing Nan, Kai Wang 외

This paper presents a novel system that enables intelligent robots to exhibit realistic body gestures while communicating with humans. The proposed system consists of a listening model and a speaking model used in corres…