paper-with-me

Papers

Research on several key technologies in practical speech emotion recognition

2017-09-27 · Chengwei Huang

In this dissertation the practical speech emotion recognition technology is studied, including several cognitive related emotion types, namely fidgetiness, confidence and tiredness. The high quality of naturalistic emotional speech data is the basis of this research. The following techniques are used for inducing practical emotional speech: cognitive task, computer game, noise stimulation, sleep deprivation and movie clips. A practical speech emotion recognition system is studied based on Gaussian mixture model. A two-class classifier set is adopted for performance improvement under the small sample case. Considering the context information in continuous emotional speech, a Gaussian mixture model embedded with Markov networks is proposed. A further study is carried out for system robustness analysis. First, noise reduction algorithm based on auditory masking properties is fist introduced to the practical speech emotion recognition. Second, to deal with the complicated unknown emotion types under real situation, an emotion recognition method with rejection ability is proposed, which enhanced the system compatibility against unknown emotion samples. Third, coping with the difficulties brought by a large number of unknown speakers, an emotional feature normalization method based on speaker-sensitive feature clustering is proposed. Fourth, by adding the electrocardiogram channel, a bi-modal emotion recognition system based on speech signals and electrocardiogram signals is first introduced. The speech emotion recognition methods studied in this dissertation may be extended into the cross-language speech emotion recognition and the whispered speech emotion recognition.

📄 PDF Abstract BibTeX arXiv:1709.09364

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringEmotion RecognitionSpeech Emotion Recognition

Similar Papers 제목 키워드 기반

Speech and Text-Based Emotion Recognizer

2023-12-10 · Varun Sharma

Affective computing is a field of study that focuses on developing systems and technologies that can understand, interpret, and respond to human emotions. Speech Emotion Recognition (SER), in particular, has got a lot of…

Data AugmentationEmotion RecognitionSpeech Emotion Recognition

CG-MER: A Card Game-based Multimodal dataset for Emotion Recognition

2025-01-14 · Nessrine Farhat, Amine Bohi, Leila Ben Letaifa, Rim Slama

The field of affective computing has seen significant advancements in exploring the relationship between emotions and emerging technologies. This paper presents a novel and valuable contribution to this field with the in…

Emotion Recognition

SyntAct: A Synthesized Database of Basic Emotions

2022-06-01 · DCLRL (LREC) 2022 6 · Felix Burkhardt, Florian Eyben, Björn Schuller

Speech emotion recognition is in the focus of research since several decades and has many applications. One problem is sparse data for supervised learning. One way to tackle this problem is the synthesis of data with emo…

Emotion RecognitionSpeech Emotion RecognitionSpeech Synthesis

Odyssey 2024 - Speech Emotion Recognition Challenge: Dataset, Baseline Framework, and Results

2024-06-20 · Odyssey: The Speaker and Language Recognition Workshop 2024 6 · Lucas Goncalves, Ali N. Salman, Abinay R. Naini, Laureano Moro Velazquez 외

The Odyssey 2024 Speech Emotion Recognition (SER) Challenge aims to enhance innovation in recognizing emotions from spontaneous speech, moving beyond traditional datasets derived from acted scenarios. It offers speaker-i…

AttributeEmotion RecognitionSpeech Emotion Recognition

New Challenges for Content Privacy in Speech and Audio

2023-01-21 · Jennifer Williams, Karla Pizzi, Shuvayanti Das, Paul-Gauthier Noe

Privacy in speech and audio has many facets. A particularly under-developed area of privacy in this domain involves consideration for information related to content and context. Speech content can include words and their…