paper-with-me

Papers

ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model

2025-10-06 · Luo Cheng, Song Siyang, Yan Siyuan, Yu Zhen, Ge Zongyuan arxiv

The automatic generation of diverse and human-like facial reactions in dyadic dialogue remains a critical challenge for human-computer interaction systems. Existing methods fail to model the stochasticity and dynamics inherent in real human reactions. To address this, we propose ReactDiff, a novel temporal diffusion framework for generating diverse facial reactions that are appropriate for responding to any given dialogue context. Our key insight is that plausible human reactions demonstrate smoothness, and coherence over time, and conform to constraints imposed by human facial anatomy. To achieve this, ReactDiff incorporates two vital priors (spatio-temporal facial kinematics) into the diffusion process: i) temporal facial behavioral kinematics and ii) facial action unit dependencies. These two constraints guide the model toward realistic human reaction manifolds, avoiding visually unrealistic jitters, unstable transitions, unnatural expressions, and other artifacts. Extensive experiments on the REACT2024 dataset demonstrate that our approach not only achieves state-of-the-art reaction quality but also excels in diversity and reaction appropriateness.

📄 PDF Abstract BibTeX arXiv:2510.04712

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ReactDiff: Latent Diffusion for Facial Reaction Generation

2025-05-20 · Jiaming Li, Sheng Wang, Xin Wang, Yitao Zhu 외

Given the audio-visual clip of the speaker, facial reaction generation aims to predict the listener's facial reactions. The challenge lies in capturing the relevance between video and audio while balancing appropriatenes…

DecoderDiversity

Reversible Graph Neural Network-based Reaction Distribution Learning for Multiple Appropriate Facial Reactions Generation

2023-05-24 · Tong Xu, Micol Spitale, Hao Tang, Lu Liu 외

Generating facial reactions in a human-human dyadic interaction is complex and highly dependent on the context since more than one facial reactions can be appropriate for the speaker's behaviour. This has challenged exis…

Graph Neural Network

ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions

2023-05-25 · Cheng Luo, Siyang Song, Weicheng Xie, Micol Spitale 외

In dyadic interaction, predicting the listener's facial reactions is challenging as different reactions could be appropriate in response to the same speaker's behaviour. Previous approaches predominantly treated this tas…

REACT2023: the first Multi-modal Multiple Appropriate Facial Reaction Generation Challenge

2023-06-11 · Siyang Song, Micol Spitale, Cheng Luo, German Barquero 외

The Multi-modal Multiple Appropriate Facial Reaction Generation Challenge (REACT2023) is the first competition event focused on evaluating multimedia processing and machine learning techniques for generating human-approp…

REACT 2024: the Second Multiple Appropriate Facial Reaction Generation Challenge

2024-01-10 · Siyang Song, Micol Spitale, Cheng Luo, Cristina Palmero 외

In dyadic interactions, humans communicate their intentions and state of mind using verbal and non-verbal cues, where multiple different facial reactions might be appropriate in response to a specific speaker behaviour. …