paper-with-me

Papers

TalkingEyes: Pluralistic Speech-Driven 3D Eye Gaze Animation

2025-01-17 · Yixiang Zhuang, Chunshan Ma, Yao Cheng, Xuan Cheng, Jing Liao, Juncong Lin

Although significant progress has been made in the field of speech-driven 3D facial animation recently, the speech-driven animation of an indispensable facial component, eye gaze, has been overlooked by recent research. This is primarily due to the weak correlation between speech and eye gaze, as well as the scarcity of audio-gaze data, making it very challenging to generate 3D eye gaze motion from speech alone. In this paper, we propose a novel data-driven method which can generate diverse 3D eye gaze motions in harmony with the speech. To achieve this, we firstly construct an audio-gaze dataset that contains about 14 hours of audio-mesh sequences featuring high-quality eye gaze motion, head motion and facial motion simultaneously. The motion data is acquired by performing lightweight eye gaze fitting and face reconstruction on videos from existing audio-visual datasets. We then tailor a novel speech-to-motion translation framework in which the head motions and eye gaze motions are jointly generated from speech but are modeled in two separate latent spaces. This design stems from the physiological knowledge that the rotation range of eyeballs is less than that of head. Through mapping the speech embedding into the two latent spaces, the difficulty in modeling the weak correlation between speech and non-verbal motion is thus attenuated. Finally, our TalkingEyes, integrated with a speech-driven 3D facial motion generator, can synthesize eye gaze motion, eye blinks, head motion and facial motion collectively from speech. Extensive quantitative and qualitative evaluations demonstrate the superiority of the proposed method in generating diverse and natural 3D eye gaze motions from speech. The project page of this paper is: https://lkjkjoiuiu.github.io/TalkingEyes_Home/

📄 PDF Abstract BibTeX arXiv:2501.09921

Code (0)

등록된 구현이 없습니다.

Tasks

Face Reconstruction

Similar Papers 제목 키워드 기반

Audio- and Gaze-driven Facial Animation of Codec Avatars

2020-08-11 · Alexander Richard, Colin Lea, Shugao Ma, Juergen Gall 외

Codec Avatars are a recent class of learned, photorealistic face models that accurately represent the geometry and texture of a person in 3D (i.e., for virtual reality), and are almost indistinguishable from video. In th…

Speech Driven Tongue Animation

2022-01-01 · CVPR 2022 1 · Salvador Medina, Denis Tome, Carsten Stoll, Mark Tiede 외

Advances in speech driven animation techniques allow the creation of convincing animations for virtual characters solely from audio data. Many existing approaches focus on facial and lip motion and they often do not …

Decoder

Unsupervised High-Resolution Portrait Gaze Correction and Animation

2022-07-01 · Jichao Zhang, Jingjing Chen, Hao Tang, Enver Sangineto 외

This paper proposes a gaze correction and animation method for high-resolution, unconstrained portrait images, which can be trained without the gaze angle and the head pose annotations. Common gaze-correction methods usu…

Image InpaintingVocal Bursts Intensity Prediction

Dual In-painting Model for Unsupervised Gaze Correction and Animation in the Wild

2020-08-09 · Jichao Zhang, Jingjing Chen, Hao Tang, Wei Wang 외

In this paper we address the problem of unsupervised gaze correction in the wild, presenting a solution that works without the need for precise annotations of the gaze angle and the head pose. We have created a new datas…

Personality-Driven Gaze Animation with Conditional Generative Adversarial Networks

2020-11-11 · Funda Durupinar

We present a generative adversarial learning approach to synthesize gaze behavior of a given personality. We train the model using an existing data set that comprises eye-tracking data and personality traits of 42 partic…

Time SeriesTime Series Analysis