paper-with-me

Papers

Latent Behavior Diffusion for Sequential Reaction Generation in Dyadic Setting

2025-05-12 · Minh-Duc Nguyen, Hyung-Jeong Yang, Soo-Hyung Kim, Ji-Eun Shin, Seung-Won Kim

The dyadic reaction generation task involves synthesizing responsive facial reactions that align closely with the behaviors of a conversational partner, enhancing the naturalness and effectiveness of human-like interaction simulations. This paper introduces a novel approach, the Latent Behavior Diffusion Model, comprising a context-aware autoencoder and a diffusion-based conditional generator that addresses the challenge of generating diverse and contextually relevant facial reactions from input speaker behaviors. The autoencoder compresses high-dimensional input features, capturing dynamic patterns in listener reactions while condensing complex input data into a concise latent representation, facilitating more expressive and contextually appropriate reaction synthesis. The diffusion-based conditional generator operates on the latent space generated by the autoencoder to predict realistic facial reactions in a non-autoregressive manner. This approach allows for generating diverse facial reactions that reflect subtle variations in conversational cues and emotional states. Experimental results demonstrate the effectiveness of our approach in achieving superior performance in dyadic reaction synthesis tasks compared to existing methods.

📄 PDF Abstract BibTeX arXiv:2505.07901

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

ReactDiff: Latent Diffusion for Facial Reaction Generation

2025-05-20 · Jiaming Li, Sheng Wang, Xin Wang, Yitao Zhu 외

Given the audio-visual clip of the speaker, facial reaction generation aims to predict the listener's facial reactions. The challenge lies in capturing the relevance between video and audio while balancing appropriatenes…

DecoderDiversity

Bayesian Inference for Jump-Diffusion Approximations of Biochemical Reaction Networks

2023-04-13 · Derya Altıntan, Bastian Alt, Heinz Koeppl

Biochemical reaction networks are an amalgamation of reactions where each reaction represents the interaction of different species. Generally, these networks exhibit a multi-scale behavior caused by the high variability …

Bayesian Inference

SynLaD: Latent Diffusion for Generating Synthesizable Molecules Conditioned on 3D Pharmacophore Profiles

2026-07-01 · Miruna Cretu, John Bradshaw, Patricia Suriana, Saeed Saremi 외 arxiv

We present SynLaD, a latent diffusion framework for small-molecule generation that unifies ligand-based drug design objectives (what to make) with synthetic accessibility (how to make it). Current models typically optimi…

From Agnostic to Specific: Latent Preference Diffusion for Multi-Behavior Sequential Recommendation

2026-02-26 · Ruochen Yang, Xiaodong Li, Jiawei Sheng, Jiangxia Cao 외 arxiv

Multi-behavior sequential recommendation (MBSR) aims to learn the dynamic and heterogeneous interactions of users' multi-behavior sequences, so as to capture user preferences under target behavior for the next interacted…

Sequential Recommendation

Non-autoregressive electron flow generation for reaction prediction

2020-12-16 · Hangrui Bi, Hengyi Wang, Chence Shi, Jian Tang

Reaction prediction is a fundamental problem in computational chemistry. Existing approaches typically generate a chemical reaction by sampling tokens or graph edits sequentially, conditioning on previously generated out…

Computational chemistryDecoderPrediction