paper-with-me

Papers

Real-Time Synchronized Interaction Framework for Emotion-Aware Humanoid Robots

2026-01-24 · Yanrong Chen, Xihan Bian arxiv

As humanoid robots increasingly introduced into social scene, achieving emotionally synchronized multimodal interaction remains a significant challenges. To facilitate the further adoption and integration of humanoid robots into service roles, we present a real-time framework for NAO robots that synchronizes speech prosody with full-body gestures through three key innovations: (1) A dual-channel emotion engine where large language model (LLM) simultaneously generates context-aware text responses and biomechanically feasible motion descriptors, constrained by a structured joint movement library; (2) Duration-aware dynamic time warping for precise temporal alignment of speech output and kinematic motion keyframes; (3) Closed-loop feasibility verification ensuring gestures adhere to NAO's physical joint limits through real-time adaptation. Evaluations show 21% higher emotional alignment compared to rule-based systems, achieved by coordinating vocal pitch (arousal-driven) with upper-limb kinematics while maintaining lower-body stability. By enabling seamless sensorimotor coordination, this framework advances the deployment of context-aware social robots in dynamic applications such as personalized healthcare, interactive education, and responsive customer service platforms.

📄 PDF Abstract BibTeX arXiv:2601.17287

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CapTalk: Text-Guided Stylization and Speech-Driven 3D Head Animation

2026-05-28 · Xuangeng Chu, Yuan Gan, Ziteng Cui, Shuhong Liu 외 arxiv

Audio-driven 3D facial animation aims to generate synchronized lip movements and vivid facial expressions from arbitrary audio clips. While existing methods can produce synchronized lip motions, they often rely on predef…

Talking Head Generation

Evaluation of Generative Models for Emotional 3D Animation Generation in VR

2025-12-18 · Kiran Chhatre, Renan Guarese, Andrii Matviienko, Christopher Peters arxiv

Social interactions incorporate nonverbal signals to convey emotions alongside speech, including facial expressions and body gestures. Generative models have demonstrated promising results in creating full-body nonverbal…

Mind-to-Face: Neural-Driven Photorealistic Avatar Synthesis via EEG Decoding

2025-12-03 · Haolin Xiong, Tianwen Fu, Pratusha Bhuvana Prasad, Yunxuan Cai 외 arxiv

Current expressive avatar systems rely heavily on visual cues, failing when faces are occluded or when emotions remain internal. We present Mind-to-Face, the first framework that decodes non-invasive electroencephalogram…

Eeg Decoding

SynchroRaMa : Lip-Synchronized and Emotion-Aware Talking Face Generation via Multi-Modal Emotion Embedding

2025-09-24 · Phyo Thet Yee, Dimitrios Kollias, Sudeepta Mishra, Abhinav Dhall arxiv

Audio-driven talking face generation has received growing interest, particularly for applications requiring expressive and natural human-avatar interaction. However, most existing emotion-aware methods rely on a single m…

Talking Face GenerationEmotion RecognitionSentiment Analysis

GaussianEmoTalker: Real-Time Emotional Talking Head Synthesis with Audio-Driven and Blendshape-Based 3D Gaussian Splatting

2026-07-01 · Haijie Yang, Zhenyu Zhang, Yixuan Dong, Jianjun Qian 외 arxiv

Audio-driven talking head synthesis has achieved impressive progress in lip synchronization and visual quality, yet generating expressive emotional avatars with controllable intensity remains challenging, especially unde…