paper-with-me

Papers

Low-Resource Adaptation for Personalized Co-Speech Gesture Generation

2022-01-01 · CVPR 2022 1 · Chaitanya Ahuja, Dong Won Lee, Louis-Philippe Morency

Personalizing an avatar for co-speech gesture generation from spoken language requires learning the idiosyncrasies of a person's gesture style from a small amount of data. Previous methods in gesture generation require large amounts of data for each speaker, which is often infeasible. We propose an approach, named DiffGAN, that efficiently personalizes co-speech gesture generation models of a high-resource source speaker to target speaker with just 2 minutes of target training data. A unique characteristic of DiffGAN is its ability to account for the crossmodal grounding shift, while also addressing the distribution shift in the output domain. We substantiate the effectiveness of our approach a large scale publicly available dataset through quantitative, qualitative and user studies, which show that our proposed methodology significantly outperforms prior approaches for low-resource adaptation of gesture generation. Code and videos can be found at https://chahuja.com/diffgan

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Gesture Generation

Similar Papers 제목 키워드 기반

Continual Learning for Personalized Co-speech Gesture Generation

2023-01-01 · ICCV 2023 1 · Chaitanya Ahuja, Pratik Joshi, Ryo Ishii, Louis-Philippe Morency

Co-speech gestures are a key channel of human communication, making them important for personalized chat agents to generate. In the past, gesture generation models assumed that data for each speaker is available all …

Continual LearningGesture Generation

Audio-Driven Co-Speech Gesture Video Generation

2022-12-05 · Xian Liu, Qianyi Wu, Hang Zhou, Yuanqi Du 외

Co-speech gesture is crucial for human-machine interaction and digital entertainment. While previous works mostly map speech audio to human skeletons (e.g., 2D keypoints), directly generating speakers' gestures in the im…

Video Generation

HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation

2025-03-03 · CVPR 2025 1 · Hongye Cheng, Tianyu Wang, Guangsi Shi, Zexing Zhao 외

Co-speech gestures are crucial non-verbal cues that enhance speech clarity and expressiveness in human communication, which have attracted increasing attention in multimodal research. While the existing methods have made…

Gesture GenerationRhythm

Super Star: Towards Streaming Real-time Interactive Agents for Digital Humans

2026-07-22 · Wentao Jiang, Youchen Xie, Haidi Fan, Yajing Chen 외 hf

Existing co-speech gesture generation methods are predominantly studied in offline settings, where gestures are synthesized from complete speech segments. However, interactive digital humans in real-world scenarios are r…

Speech-driven Personalized Gesture Synthetics: Harnessing Automatic Fuzzy Feature Inference

2024-03-16 · Fan Zhang, Zhaohan Wang, Xin Lyu, Siyuan Zhao 외

Speech-driven gesture generation is an emerging field within virtual human creation. However, a significant challenge lies in accurately determining and processing the multitude of input features (such as acoustic, seman…

Gesture Generation