paper-with-me

Papers

Efficient Emotional Adaptation for Audio-Driven Talking-Head Generation

2023-09-10 · ICCV 2023 1 · Yuan Gan, Zongxin Yang, Xihang Yue, Lingyun Sun, Yi Yang

Audio-driven talking-head synthesis is a popular research topic for virtual human-related applications. However, the inflexibility and inefficiency of existing methods, which necessitate expensive end-to-end training to transfer emotions from guidance videos to talking-head predictions, are significant limitations. In this work, we propose the Emotional Adaptation for Audio-driven Talking-head (EAT) method, which transforms emotion-agnostic talking-head models into emotion-controllable ones in a cost-effective and efficient manner through parameter-efficient adaptations. Our approach utilizes a pretrained emotion-agnostic talking-head transformer and introduces three lightweight adaptations (the Deep Emotional Prompts, Emotional Deformation Network, and Emotional Adaptation Module) from different perspectives to enable precise and realistic emotion controls. Our experiments demonstrate that our approach achieves state-of-the-art performance on widely-used benchmarks, including LRW and MEAD. Additionally, our parameter-efficient adaptations exhibit remarkable generalization ability, even in scenarios where emotional training videos are scarce or nonexistent. Project website: https://yuangan.github.io/eat/

📄 PDF Abstract BibTeX arXiv:2309.04946

Code (1)

yuangan/eat_code 공식 구현 pytorch

Tasks

Talking Head Generation

Similar Papers 제목 키워드 기반

EmoGene: Audio-Driven Emotional 3D Talking-Head Generation

2024-10-07 · Wenqing Wang, Yun Fu

Audio-driven talking-head generation is a crucial and useful technology for virtual human interaction and film-making. While recent advances have focused on improving image fidelity and lip synchronization, generating ac…

NeRFTalking Head Generation

GaussianEmoTalker: Real-Time Emotional Talking Head Synthesis with Audio-Driven and Blendshape-Based 3D Gaussian Splatting

2026-07-01 · Haijie Yang, Zhenyu Zhang, Yixuan Dong, Jianjun Qian 외 arxiv

Audio-driven talking head synthesis has achieved impressive progress in lip synchronization and visual quality, yet generating expressive emotional avatars with controllable intensity remains challenging, especially unde…

Emotional Talking Head Generation based on Memory-Sharing and Attention-Augmented Networks

2023-06-06 · Jianrong Wang, Yaxin Zhao, Li Liu, Tianyi Xu 외

Given an audio clip and a reference face image, the goal of the talking head generation is to generate a high-fidelity talking head video. Although some audio-driven methods of generating talking head videos have made so…

Talking Head Generation

EMOdiffhead: Continuously Emotional Control in Talking Head Generation via Diffusion

2024-09-11 · Jian Zhang, Weijian Mai, Zhijun Zhang

The task of audio-driven portrait animation involves generating a talking head video using an identity image and an audio track of speech. While many existing approaches focus on lip synchronization and video quality, fe…

Portrait AnimationTalking Head GenerationVideo Generation

EmoCAST: Emotional Talking Portrait via Emotive Text Description

2025-08-28 · Yiguo Jiang, Xiaodong Cun, Yong Zhang, Yudian Zheng 외 arxiv

Emotional talking head synthesis aims to generate talking portrait videos with vivid expressions. Existing methods still exhibit limitations in control flexibility, motion naturalness, and expression quality. Moreover, c…

Motion Synthesis