paper-with-me

홈 › Papers

A neural network based post-filter for speech-driven head motion synthesis

2019-07-24 · JinHong Lu, Hiroshi Shimodaira

Despite the fact that neural networks are widely used for speech-driven head motion synthesis, it is well-known that the output of neural networks is noisy or discontinuous due to the limited capability of deep neural networks in predicting human motion. Thus, post-processing is required to obtain smooth head motion trajectories for animation. It is common to apply a linear filter or consider keyframes as post-processing. However, neither approach is optimal as there is always a trade-off between smoothness and accuracy. We propose to employ a neural network trained in a way that it is capable of reconstructing the head motions, in order to overcome this limitation. In the objective evaluation, this filter is proved to be good at de-noising data involving types of noise (dropout or Gaussian noise). Objective metrics also demonstrate the improvement of the joined head motion's smoothness after being processed by our proposed filter. A detailed analysis reveals that our proposed filter learns the characteristic of head motions. The subjective evaluation shows that participants were unable to distinguish the synthesised head motions with our proposed filter from ground truth, which was preferred over the Gaussian filter and moving average.

📄 PDF Abstract BibTeX arXiv:1907.10585

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Synthesis

Similar Papers 제목 키워드 기반

SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation

2024-05-12 · Changpeng Cai, Guinan Guo, Jiao Li, Junhao Su 외

Most earlier researches on talking face generation have focused on the synchronization of lip motion and speech content. However, head pose and facial emotions are equally important characteristics of natural faces. Whil…

DisentanglementFace GenerationTalking Face GenerationTalking Head Generation

TalkingEyes: Pluralistic Speech-Driven 3D Eye Gaze Animation

2025-01-17 · Yixiang Zhuang, Chunshan Ma, Yao Cheng, Xuan Cheng 외

Although significant progress has been made in the field of speech-driven 3D facial animation recently, the speech-driven animation of an indispensable facial component, eye gaze, has been overlooked by recent research. …

Face Reconstruction

Enhancement Of Coded Speech Using a Mask-Based Post-Filter

2020-10-12 · Srikanth Korse, Kishan Gupta, Guillaume Fuchs

The quality of speech codecs deteriorates at low bitrates due to high quantization noise. A post-filter is generally employed to enhance the quality of the coded speech. In this paper, a data-driven post-filter relying o…

DecoderQuantization

Moving fast and slow: Analysis of representations and post-processing in speech-driven automatic gesture generation

2020-07-16 · Taras Kucherenko, Dai Hasegawa, Naoshi Kaneko, Gustav Eje Henter 외

This paper presents a novel framework for speech-driven gesture production, applicable to virtual agents to enhance human-computer interaction. Specifically, we extend recent deep-learning-based, data-driven methods for …

Gesture GenerationRepresentation Learning

Transformer Network for Semantically-Aware and Speech-Driven Upper-Face Generation

2021-10-09 · Mireille Fares, Catherine Pelachaud, Nicolas Obin

We propose a semantically-aware speech driven model to generate expressive and natural upper-facial and head motion for Embodied Conversational Agents (ECA). In this work, we aim to produce natural and continuous head mo…

Face Generation