Laughter Synthesis: Combining Seq2seq modeling with Transfer Learning
Despite the growing interest for expressive speech synthesis, synthesis of nonverbal expressions is an under-explored area. In this paper we propose an audio laughter synthesis system based on a sequence-to-sequence TTS synthesis system. We leverage transfer learning by training a deep learning model to learn to generate both speech and laughs from annotations. We evaluate our model with a listening test, comparing its performance to an HMM-based laughter synthesis one and assess that it reaches higher perceived naturalness. Our solution is a first step towards a TTS system that would be able to synthesize speech with a control on amusement level with laughter integration.
Code (1)
Tasks
Expressive Speech SynthesisSpeech SynthesisTransfer LearningSimilar Papers 제목 키워드 기반
Generating Diverse Realistic Laughter for Interactive Art
We propose an interactive art project to make those rendered invisible by the COVID-19 crisis and its concomitant solitude reappear through the welcome melody of laughter, and connections created and explored through adv…
DiversityThe AV-LASYN Database : A synchronous corpus of audio and 3D facial marker data for audio-visual laughter synthesis
A synchronous database of acoustic and 3D facial marker data was built for audio-visual laughter synthesis. Since the aim is to use this database for HMM-based modeling and synthesis, the amount of collected data from on…
Dimensionality ReductionSpeech SynthesisA generative framework for conversational laughter: Its 'language model' and laughter sound synthesis
As the phonetic and acoustic manifestations of laughter in conversation are highly diverse, laughter synthesis should be capable of accommodating such diversity while maintaining high controllability. This paper proposes…
DiversityLanguage ModelingLanguage ModellingLaughNet: synthesizing laughter utterances from waveform silhouettes and a single laughter example
Emotional and controllable speech synthesis is a topic that has received much attention. However, most studies focused on improving the expressiveness and controllability in the context of linguistic content, even though…
Speech SynthesisExecutive Voiced Laughter and Social Approval: An Explorative Machine Learning Study
We study voiced laughter in executive communication and its effect on social approval. Integrating research on laughter, affect-as-information, and infomediaries' social evaluations of firms, we hypothesize that voiced l…
Sentiment Analysis