GestureGAN for Hand Gesture-to-Gesture Translation in the Wild
Hand gesture-to-gesture translation in the wild is a challenging task since hand gestures can have arbitrary poses, sizes, locations and self-occlusions. Therefore, this task requires a high-level understanding of the mapping between the input source gesture and the output target gesture. To tackle this problem, we propose a novel hand Gesture Generative Adversarial Network (GestureGAN). GestureGAN consists of a single generator $G$ and a discriminator $D$, which takes as input a conditional hand image and a target hand skeleton image. GestureGAN utilizes the hand skeleton information explicitly, and learns the gesture-to-gesture mapping through two novel losses, the color loss and the cycle-consistency loss. The proposed color loss handles the issue of "channel pollution" while back-propagating the gradients. In addition, we present the Fr\'echet ResNet Distance (FRD) to evaluate the quality of generated images. Extensive experiments on two widely used benchmark datasets demonstrate that the proposed GestureGAN achieves state-of-the-art performance on the unconstrained hand gesture-to-gesture translation task. Meanwhile, the generated images are in high-quality and are photo-realistic, allowing them to be used as data augmentation to improve the performance of a hand gesture classifier. Our model and code are available at https://github.com/Ha0Tang/GestureGAN.
Code (1)
Tasks
Data AugmentationGenerative Adversarial NetworkGesture-to-Gesture TranslationTranslationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Gesture-to-Gesture Translation in the Wild via Category-Independent Conditional Maps
Recent works have shown Generative Adversarial Networks (GANs) to be particularly effective in image-to-image translations. However, in tasks such as body pose and hand gesture translation, existing methods usually requi…
Gesture-to-Gesture TranslationTranslationUnified Generative Adversarial Networks for Controllable Image-to-Image Translation
We propose a unified Generative Adversarial Network (GAN) for controllable image-to-image translation, i.e., transferring an image from a source to a target domain guided by controllable structures. In addition to condit…
Facial Expression TranslationGenerative Adversarial NetworkGesture-to-Gesture TranslationImage Generation+2Learning Individual Styles of Conversational Gesture
Human speech is often accompanied by hand and arm gestures. Given audio speech input, we generate plausible gestures to go along with the sound. Specifically, we perform cross-modal translation from "in-the-wild'' monolo…
Gesture GenerationSpeech-to-Gesture TranslationTranslationSHREC 2021: Track on Skeleton-based Hand Gesture Recognition in the Wild
Gesture recognition is a fundamental tool to enable novel interaction paradigms in a variety of application scenarios like Mixed Reality environments, touchless public kiosks, entertainment systems, and more. Recognition…
Action RecognitionGesture RecognitionHand Gesture RecognitionHand-Gesture Recognition+1Fast and Robust Dynamic Hand Gesture Recognition via Key Frames Extraction and Feature Fusion
Gesture recognition is a hot topic in computer vision and pattern recognition, which plays a vitally important role in natural human-computer interface. Although great progress has been made recently, fast and robust han…
ClusteringGesture RecognitionHand Gesture RecognitionHand-Gesture Recognition