paper-with-me

Papers

Gesture2Vec: Clustering Gestures using Representation Learning Methods for Co-speech Gesture Generation

2021-09-29 · Payam Jome Yazdian, Mo Chen, Angelica Lim

Co-speech gestures are a principal component in conveying messages and enhancing interaction experiences between humans. Similarly, the co-speech gesture is a key ingredient in human-agent interaction including both virtual agents and robots. Existing machine learning approaches have yielded only marginal success in learning speech-to-motion at the frame level. Current methods generate repetitive gesture sequences that lack appropriateness with respect to the speech context. In this paper, we propose a Gesture2Vec model using representation learning methods to learn the relationship between semantic features and corresponding gestures. We propose a vector-quantized variational autoencoder structure as well as training techniques to learn a rigorous representation of gesture sequences. Furthermore, we use a machine translation model that takes input text and translates it into a discrete sequence of associated gesture chunks in the learned gesture space. Ultimately, we use translated quantized gestures from the input text as an input to the autoencoder’s decoder to produce gesture sequences. The resulting gestures can be applied to both virtual agents and humanoid robots. Subjective and objective evaluations confirm the success of our approach in terms of appropriateness, human-likeness, and diversity.

📄 PDF Abstract BibTeX

Code (1)

pjyazdian/Gesture2Vec 공식 구현 pytorch

Tasks

ClusteringDecoderDiversityGesture GenerationMachine TranslationRepresentation LearningTranslation

Similar Papers 제목 키워드 기반

Speech2Properties2Gestures: Gesture-Property Prediction as a Tool for Generating Representational Gestures from Speech

2021-06-28 · Taras Kucherenko, Rajmund Nagy, Patrik Jonell, Michael Neff 외

We propose a new framework for gesture generation, aiming to allow data-driven approaches to produce more semantically rich gestures. Our approach first predicts whether to gesture, followed by a prediction of the gestur…

Gesture GenerationProperty Prediction

Learning Co-Speech Gesture Representations in Dialogue through Contrastive Learning: An Intrinsic Evaluation

2024-08-31 · Esam Ghaleb, Bulat Khaertdinov, Wim Pouw, Marlou Rasenberg 외

In face-to-face dialogues, the form-meaning relationship of co-speech gestures varies depending on contextual factors such as what the gestures refer to and the individual characteristics of speakers. These factors make …

Contrastive LearningDiagnosticRepresentation Learning

Learning Hierarchical Cross-Modal Association for Co-Speech Gesture Generation

2022-03-24 · CVPR 2022 1 · Xian Liu, Qianyi Wu, Hang Zhou, Yinghao Xu 외

Generating speech-consistent body and gesture movements is a long-standing problem in virtual avatar creation. Previous studies often synthesize pose movement in a holistic manner, where poses of all joints are generated…

Contrastive LearningGesture Generation

Understanding Co-speech Gestures in-the-wild

2025-03-28 · Sindhu B Hegde, K R Prajwal, Taein Kwon, Andrew Zisserman

Co-speech gestures play a vital role in non-verbal communication. In this paper, we introduce a new framework for co-speech gesture understanding in the wild. Specifically, we propose three new tasks and benchmarks to ev…

Active Speaker Detection

Investigating the impact of 2D gesture representation on co-speech gesture generation

2024-06-21 · Teo Guichoux, Laure Soulier, Nicolas Obin, Catherine Pelachaud

Co-speech gestures play a crucial role in the interactions between humans and embodied conversational agents (ECA). Recent deep learning methods enable the generation of realistic, natural co-speech gestures synchronized…

3D Pose EstimationGesture GenerationPose Estimation