paper-with-me

홈 › Papers

VOCAL: Vowel and Consonant Layering for Expressive Animator-Centric Singing Animation

2022-11-30 · Siggraph Asia 2022 2022 11 · Yifang Pan, Chris Landreth, Eugene Fiume, Karan Singh Authors Info & Claims

Singing and speaking are two fundamental forms of human communication. From a modeling perspective however, speaking can be seen as a subset of singing. We present VOCAL, a system that automatically generates expressive, animator-centric lower face animation from singing audio input. Articulatory phonetics and voice instruction ascribe additional roles to vowels (projecting melody and volume) and consonants (lyrical clarity and rhythmic emphasis) in song. Our approach directly uses these insights to define axes for Melodic-accent and Pitch-sensitivity (Ma-Ps), which together provide an abstract space to visually represent various singing styles. In our system. vowels are processed first. A lyrical vowel is often sung tonally as one or more different vowels. We perform any such vowel modifications using a neural network trained on input audio. These vowels are then dilated from their spoken behaviour to bleed into each other based on Melodic-accent (Ma), with Pitch-sensitivity (Ps) modeling visual vibrato. Consonant animation curves are then layered in, with viseme intensity modeling rhythmic emphasis (inverse to Ma). Our evaluation is fourfold: we show the impact of our design parameters; we compare our results to ground truth and prior art; we present compelling results on a variety of voices and singing styles; and we validate these results with professional singers and animators.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Sensitivity

Similar Papers 제목 키워드 기반

Vowel and Consonant Classification through Spectral Decomposition

2017-09-01 · WS 2017 9 · Patricia Thaine, Gerald Penn

We consider two related problems in this paper. Given an undeciphered alphabetic writing system or mono-alphabetic cipher, determine: (1) which of its letters are vowels and which are consonants; and (2) whether the writ…

ClassificationGeneral Classification

Transmission of droplet-conveyed infectious agents such as SARS-CoV-2 by speech and vocal exercises during speech therapy: preliminary experiment concerning airflow velocity

2021-01-06 · Antoine Giovanni, Thomas Radulesco, Gilles Bouchet, Alexia Mattei 외

Purpose Infectious agents, such as SARS-CoV-2, can be carried by droplets expelled during breathing. The spatial dissemination of droplets varies according to their initial velocity. After a short literature review, our …

Employing self-supervised learning models for cross-linguistic child speech maturity classification

2025-06-10 · Theo Zhang, Madurya Suresh, Anne S. Warlaumont, Kasia Hitczenko 외

Speech technology systems struggle with many downstream tasks for child speech due to small training corpora and the difficulties that child speech pose. We apply a novel dataset, SpeechMaturity, to state-of-the-art tran…

Self-Supervised Learningvalid

Articulatory modeling of the S-shaped F2 trajectories observed in Öhman's spectrographic analysis of VCV syllables

2025-05-28 · Frédéric Berthommier

The synthesis of Ohman's VCV sequences with intervocalic plosive consonants was first achieved 30 years ago using the DRM model. However, this approach remains primarily acoustic and lacks articulatory constraints. In th…

Trajectory Planning

Modeling and Estimation of Vocal Tract and Glottal Source Parameters Using ARMAX-LF Model

2024-10-07 · Kai Lia, Masato Akagia, Yongwei Lib, Masashi Unokia

Modeling and estimation of the vocal tract and glottal source parameters of vowels from raw speech can be typically done by using the Auto-Regressive with eXogenous input (ARX) model and Liljencrants-Fant (LF) model with…