paper-with-me

홈 › Papers

Acoustically-Driven Phoneme Removal That Preserves Vocal Affect Cues

2022-10-26 · Camille Noufi, Jonathan Berger, Karen J. Parker, Daniel L. Bowling

In this paper, we propose a method for removing linguistic information from speech for the purpose of isolating paralinguistic indicators of affect. The immediate utility of this method lies in clinical tests of sensitivity to vocal affect that are not confounded by language, which is impaired in a variety of clinical populations. The method is based on simultaneous recordings of speech audio and electroglottographic (EGG) signals. The speech audio signal is used to estimate the average vocal tract filter response and amplitude envelop. The EGG signal supplies a direct correlate of voice source activity that is mostly independent of phonetic articulation. The dynamic energy of the speech audio and the average vocal tract filter are applied to the EGG signal create a third signal designed to capture as much paralinguistic information from the vocal production system as possible -- maximizing the retention of bioacoustic cues to affect -- while eliminating phonetic cues to verbal meaning. To evaluate the success of this method, we studied the perception of corresponding speech audio and transformed EGG signals in an affect rating experiment with online listeners. The results show a high degree of similarity in the perceived affect of matched signals, indicating that our method is effective.

📄 PDF Abstract BibTeX arXiv:2210.15001

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Singing voice synthesis based on frame-level sequence-to-sequence models considering vocal timing deviation

2023-01-05 · Miku Nishihara, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku 외

This paper proposes singing voice synthesis (SVS) based on frame-level sequence-to-sequence models considering vocal timing deviation. In SVS, it is essential to synchronize the timing of singing with temporal structures…

Singing Voice Synthesis

Phonetic and Lexical Discovery of a Canine Language using HuBERT

2024-02-25 · Xingyuan Li, Sinong Wang, Zeyu Xie, Mengyue Wu 외

This paper delves into the pioneering exploration of potential communication patterns within dog vocalizations and transcends traditional linguistic analysis barriers, which heavily relies on human priori knowledge on li…

Phoneme-Informed Note Segmentation of Monophonic Vocal Music

2021-11-01 · NLP4MusA 2021 11 · Yukun Li, Emir Demirel, Polina Proutskova, Simon Dixon

The phonetic bases of vocal expressed emotion: natural versus acted

2019-11-13 · Hira Dhamyal, Shahan Ali Memon, Bhiksha Raj, Rita Singh

Can vocal emotions be emulated? This question has been a recurrent concern of the speech community, and has also been vigorously investigated. It has been fueled further by its link to the issue of validity of acted emot…

Emotion ClassificationGeneral Classificationvalid

Arti-JEPA: Adapting Video World Model to Real-Time MRI of the Vocal Tract for Speech-Production Analysis

2026-09-09 · Hong Nguyen, Sean Foley, Christina Hagedorn, Yijing Lu 외 arxiv

Real-time MRI (rtMRI) captures the dynamics of the entire vocal tract during speech, but labeled data are scarce and the modality - single-slice, grayscale, low-resolution - differs substantially from the natural videos …

Domain Adaptation