Gesture Generation
4개 벤치마크 · 논문 137편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
Generating Holistic 3D Human Motion from Speech
The GENEA Challenge 2022: A large evaluation of data-driven co-speech gesture generation
The GENEA Challenge 2023: A large scale evaluation of gesture generation models in monadic and dyadic settings
Speech Gesture Generation from the Trimodal Context of Text, Audio, and Speaker Identity
Learning Individual Styles of Conversational Gesture
Papers
Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation
Generating co-speech gestures that are temporally coherent, semantically aligned with speech, and grounded with surrounding objects remains challenging. Prior speech-driven gesture models emphasize audio-gesture alignmen…
Gesture GenerationInteractGesture: Progressive Chunk Guidance for Continuous Streaming Co-Speech Gesture Control
Co-speech gesture generation has made significant progress toward realistic full-body motion from speaker audio, yet existing models lack fine-grained spatial controllability of individual joints. To address this, we int…
Gesture GenerationSICAGE: Speaker-Independent Culture-Aware Gesture Generation using TED4C-L Dataset
Recent co-speech gesture generation methods often overlook cultural differences, limiting their effectiveness in human-agent interaction. Moreover, culture-conditioned models are rarely evaluated under speaker-disjoint s…
Domain GeneralizationGesture GenerationMotion SynthesisGenerating Natural and Expressive Robot Gestures through Iterative Reinforcement Learning with Human Feedback using LLMs
Expressive gestures are essential for natural and effective communication, complementing speech when verbal cues alone are insufficient (e.g., pointing). For social robots such as the humanoid Pepper, producing natural a…
Reinforcement LearningGesture GenerationCode GenerationSiGnature: Explicit Motion Diffusion for Stylized Semantic Gesture
While recent advances in co-speech gesture generation have achieved impressive rhythmic synchronization, synthesizing gestures that are both semantically meaningful and faithful to a speaker's unique non-verbal style rem…
Gesture GenerationMamba-Enhanced Implicit Motion Learning for Audio-Driven Portrait Animation
Audio-driven human motion video generation aims to synthesize realistic and temporally coherent human animations from a single static image, with applications in talking-head synthesis, co-speech gesture generation, and …
Gesture GenerationVideo Generation