paper-with-me

홈 › Papers

On The Differences Between Song and Speech Emotion Recognition: Effect of Feature Sets, Feature Types, and Classifiers

2020-04-01 · Bagus Tris Atmaja, Masato Akagi

In this paper, we evaluate the different features sets, feature types, and classifiers on both song and speech emotion recognition. Three feature sets: GeMAPS, pyAudioAnalysis, and LibROSA; two feature types: low-level descriptors and high-level statistical functions; and four classifiers: multilayer perceptron, LSTM, GRU, and convolution neural networks are examined on both song and speech data with the same parameter values. The results show no remarkable difference between song and speech data using the same method. In addition, high-level statistical functions of acoustic features gained higher performance scores than low-level descriptors in this classification task. This result strengthens the previous finding on the regression task which reported the advantage use of high-level features.

📄 PDF Abstract BibTeX arXiv:2004.00200

Code (1)

bagustris/ravdess_song_speech 공식 구현 tf

Tasks

Emotion RecognitionregressionSpeech Emotion Recognition

Similar Papers 제목 키워드 기반

Speech & Song Emotion Recognition Using Multilayer Perceptron and Standard Vector Machine

2021-05-19 · Behzad Javaheri

Herein, we have compared the performance of SVM and MLP in emotion recognition using speech and song channels of the RAVDESS dataset. We have undertaken a journey to extract various audio features, identify optimal scali…

Data AugmentationEmotion Recognition

Feature Selection Enhancement and Feature Space Visualization for Speech-Based Emotion Recognition

2022-08-19 · Sofia Kanwal, Sohail Asghar, Hazrat Ali

Robust speech emotion recognition relies on the quality of the speech features. We present speech features enhancement strategy that improves speech emotion recognition. We used the INTERSPEECH 2010 challenge feature-set…

Emotion Recognitionfeature selectionSpeech Emotion Recognition

emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation

2023-12-23 · Ziyang Ma, Zhisheng Zheng, Jiaxin Ye, Jinchao Li 외

We propose emotion2vec, a universal speech emotion representation model. emotion2vec is pre-trained on open-source unlabeled emotion data through self-supervised online distillation, combining utterance-level loss and fr…

Emotion RecognitionSelf-Supervised LearningSentiment AnalysisSpeech Emotion Recognition

Emotion Recognition from Speech

2019-12-22 · Kannan Venkataramanan, Haresh Rengaraj Rajamohan

In this work, we conduct an extensive comparison of various approaches to speech based emotion recognition systems. The analyses were carried out on audio recordings from Ryerson Audio-Visual Database of Emotional Speech…

Emotion ClassificationEmotion RecognitionGeneral Classification

Speech Emotion Recognition Using Quaternion Convolutional Neural Networks

2021-10-31 · Aneesh Muppidi, Martin Radfar

Although speech recognition has become a widespread technology, inferring emotion from speech signals still remains a challenge. To address this problem, this paper proposes a quaternion convolutional neural network (QCN…

Emotion RecognitionSpeech Emotion Recognitionspeech-recognitionSpeech Recognition