paper-with-me

Papers

Audio- and Gaze-driven Facial Animation of Codec Avatars

2020-08-11 · Alexander Richard, Colin Lea, Shugao Ma, Juergen Gall, Fernando de la Torre, Yaser Sheikh

Codec Avatars are a recent class of learned, photorealistic face models that accurately represent the geometry and texture of a person in 3D (i.e., for virtual reality), and are almost indistinguishable from video. In this paper we describe the first approach to animate these parametric models in real-time which could be deployed on commodity virtual reality hardware using audio and/or eye tracking. Our goal is to display expressive conversations between individuals that exhibit important social signals such as laughter and excitement solely from latent cues in our lossy input signals. To this end we collected over 5 hours of high frame rate 3D face scans across three participants including traditional neutral speech as well as expressive and conversational speech. We investigate a multimodal fusion approach that dynamically identifies which sensor encoding should animate which parts of the face at any time. See the supplemental video which demonstrates our ability to generate full face motion far beyond the typically neutral lip articulations seen in competing work: https://research.fb.com/videos/audio-and-gaze-driven-facial-animation-of-codec-avatars/

📄 PDF Abstract BibTeX arXiv:2008.05023

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TalkingEyes: Pluralistic Speech-Driven 3D Eye Gaze Animation

2025-01-17 · Yixiang Zhuang, Chunshan Ma, Yao Cheng, Xuan Cheng 외

Although significant progress has been made in the field of speech-driven 3D facial animation recently, the speech-driven animation of an indispensable facial component, eye gaze, has been overlooked by recent research. …

Face Reconstruction

From Tokens to Faces: Investigating Discrete Speech Representations for 3D Facial Animation

2026-06-11 · Pedro Correa, Olivier Perrotin, Samir Sadok, Paula Costa 외 arxiv

The choice of speech representation is critical in speech-driven 3D facial animation. Representations differ in what they encode: SSL features emphasize segmental and semantic cues, neural codecs yield latents optimized …

Audio-Driven Talking Face Generation with Diverse yet Realistic Facial Animations

2023-04-18 · Rongliang Wu, Yingchen Yu, Fangneng Zhan, Jiahui Zhang 외

Audio-driven talking face generation, which aims to synthesize talking faces with realistic facial animations (including accurate lip movements, vivid facial expression details and natural head poses) corresponding to th…

Face GenerationTalking Face Generation

Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars

2025-08-22 · NVIDIA, :, Chaeyeon Chung, Ilya Fedorov 외 arxiv

Audio-driven facial animation presents an effective solution for animating digital avatars. In this paper, we detail the technical aspects of NVIDIA Audio2Face-3D, including data acquisition, network architecture, retarg…

Takin-ADA: Emotion Controllable Audio-Driven Animation with Canonical and Landmark Loss Optimization

2024-10-18 · Bin Lin, Yanzhen Yu, Jianhao Ye, Ruitao Lv 외

Existing audio-driven facial animation methods face critical challenges, including expression leakage, ineffective subtle expression transfer, and imprecise audio-driven synchronization. We discovered that these issues s…

GPUPortrait Animation