paper-with-me

Papers

EMOVO Corpus: an Italian Emotional Speech Database

2014-05-01 · LREC 2014 5 · Giovanni Costantini, Iacopo Iaderola, Andrea Paoloni, Massimiliano Todisco

This article describes the first emotional corpus, named EMOVO, applicable to Italian language,. It is a database built from the voices of up to 6 actors who played 14 sentences simulating 6 emotional states (disgust, fear, anger, joy, surprise, sadness) plus the neutral state. These emotions are the well-known Big Six found in most of the literature related to emotional speech. The recordings were made with professional equipment in the Fondazione Ugo Bordoni laboratories. The paper also describes a subjective validation test of the corpus, based on emotion-discrimination of two sentences carried out by two different groups of 24 listeners. The test was successful because it yielded an overall recognition accuracy of 80{\%}. It is observed that emotions less easy to recognize are joy and disgust, whereas the most easy to detect are anger, sadness and the neutral state.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion Recognition

Similar Papers 제목 키워드 기반

Transfer learning from High-Resource to Low-Resource Language Improves Speech Affect Recognition Classification Accuracy

2021-03-04 · Sara Durrani, Umair Arshad

Speech Affect Recognition is a problem of extracting emotional affects from audio data. Low resource languages corpora are rear and affect recognition is a difficult task in cross-corpus settings. We present an approach …

Cross-corpusDiversityTransfer Learning

Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages

2024-02-27 · Lucía Gómez Zaragozá, Rocío del Amor, Elena Parra Vargas, Valery Naranjo 외

Emotional Voice Messages (EMOVOME) is a spontaneous speech dataset containing 999 audio messages from real conversations on a messaging app from 100 Spanish speakers, gender balanced. Voice messages were produced in-the-…

Emotion Recognition

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition

2025-01-06 · Ruoyu Zhao, Xiantao Jiang, F. Richard Yu, Victor C. M. Leung 외

Speech Emotion Recognition (SER) plays a crucial role in enhancing human-computer interaction. Cross-Linguistic SER (CLSER) has been a challenging research problem due to significant variability in linguistic and acousti…

Emotion RecognitionSpeech Emotion RecognitionTransfer Learning

EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting

2025-04-17 · Guanrou Yang, Chen Yang, Qian Chen, Ziyang Ma 외

Human speech goes beyond the mere transfer of information; it is a profound exchange of emotions and a connection between individuals. While Text-to-Speech (TTS) models have made huge progress, they still face challenges…

text-to-speechText to Speech

EMOVOME: A Dataset for Emotion Recognition in Spontaneous Real-Life Speech

2024-03-04 · Lucía Gómez-Zaragozá, Rocío del Amor, María José Castro-Bleda, Valery Naranjo 외

Spontaneous datasets for Speech Emotion Recognition (SER) are scarce and frequently derived from laboratory environments or staged scenarios, such as TV shows, limiting their application in real-world contexts. We develo…

Emotion RecognitionFairnessSpeech Emotion Recognition