paper-with-me

홈 › Papers

Laughing Matters: Introducing Laughing-Face Generation using Diffusion Models

2023-05-15 · Antoni Bigata Casademunt, Rodrigo Mira, Nikita Drobyshev, Konstantinos Vougioukas, Stavros Petridis, Maja Pantic

Speech-driven animation has gained significant traction in recent years, with current methods achieving near-photorealistic results. However, the field remains underexplored regarding non-verbal communication despite evidence demonstrating its importance in human interaction. In particular, generating laughter sequences presents a unique challenge due to the intricacy and nuances of this behaviour. This paper aims to bridge this gap by proposing a novel model capable of generating realistic laughter sequences, given a still portrait and an audio clip containing laughter. We highlight the failure cases of traditional facial animation methods and leverage recent advances in diffusion models to produce convincing laughter videos. We train our model on a diverse set of laughter datasets and introduce an evaluation metric specifically designed for laughter. When compared with previous speech-driven approaches, our model achieves state-of-the-art performance across all metrics, even when these are re-trained for laughter generation. Our code and project are publicly available

📄 PDF Abstract BibTeX arXiv:2305.08854

Code (1)

antonibigata/Laughing-Matters 공식 구현 pytorch

Tasks

Face Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Detection of Children Abuse by Voice and Audio Classification by Short-Time Fourier Transform Machine Learning implemented on Nvidia Edge GPU device

2023-07-27 · Jiuqi Yan, Yingxian Chen, W. W. T. Fok

The safety of children in children home has become an increasing social concern, and the purpose of this experiment is to use machine learning applied to detect the scenarios of child abuse to increase the safety of chil…

Abuse DetectionAudio ClassificationGPUimage-classification+1

Inhalation Noises as Endings of Laughs in Conversational Speech

2022-06-01 · SmiLa (LREC) 2022 6 · Jürgen Trouvain, Raphael Werner, Khiet Truong

In this study we investigate the role of inhalation noises at the end of laughter events in two conversational corpora that provide relevant annotations. A re-annotation of the categories for laughter, silence and inbrea…

Self-supervised learning of a facial attribute embedding from video

2018-08-21 · Olivia Wiles, A. Sophia Koepke, Andrew Zisserman

We propose a self-supervised framework for learning facial attributes by simply watching videos of a human face speaking, laughing, and moving over time. To perform this task, we introduce a network, Facial Attributes-Ne…

AttributeSelf-Supervised LearningUnsupervised Facial Landmark Detection

Making Flow-Matching-Based Zero-Shot Text-to-Speech Laugh as You Like

2024-02-12 · Naoyuki Kanda, Xiaofei Wang, Sefik Emre Eskimez, Manthan Thakker 외

Laughter is one of the most expressive and natural aspects of human speech, conveying emotions, social cues, and humor. However, most text-to-speech (TTS) systems lack the ability to produce realistic and appropriate lau…

text-to-speechText to Speech

LaughTalk: Expressive 3D Talking Head Generation with Laughter

2023-11-02 · Kim Sung-Bin, Lee Hyun, Da Hye Hong, Suekyeong Nam 외

Laughter is a unique expression, essential to affirmative social interactions of humans. Although current 3D talking head generation methods produce convincing verbal articulations, they often fail to capture the vitalit…

Talking Head Generation