paper-with-me

홈 › Papers

Noise robust speech emotion recognition with signal-to-noise ratio adapting speech enhancement

2023-09-03 · Yu-Wen Chen, Julia Hirschberg, Yu Tsao

Speech emotion recognition (SER) often experiences reduced performance due to background noise. In addition, making a prediction on signals with only background noise could undermine user trust in the system. In this study, we propose a Noise Robust Speech Emotion Recognition system, NRSER. NRSER employs speech enhancement (SE) to effectively reduce the noise in input signals. Then, the signal-to-noise-ratio (SNR)-level detection structure and waveform reconstitution strategy are introduced to reduce the negative impact of SE on speech signals with no or little background noise. Our experimental results show that NRSER can effectively improve the noise robustness of the SER system, including preventing the system from making emotion recognition on signals consisting solely of background noise. Moreover, the proposed SNR-level detection structure can be used individually for tasks such as data selection.

📄 PDF Abstract BibTeX arXiv:2309.01164

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion RecognitionSpeech Enhancement

Similar Papers 제목 키워드 기반

Research on several key technologies in practical speech emotion recognition

2017-09-27 · Chengwei Huang

In this dissertation the practical speech emotion recognition technology is studied, including several cognitive related emotion types, namely fidgetiness, confidence and tiredness. The high quality of naturalistic emoti…

ClusteringEmotion RecognitionSpeech Emotion Recognition

TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition

2024-04-19 · Chengxin Chen, Pengyuan Zhang

One persistent challenge in Speech Emotion Recognition (SER) is the ubiquitous environmental noise, which frequently results in deteriorating SER performance in practice. In this paper, we introduce a Two-level Refinemen…

Emotion RecognitionSpeech Emotion RecognitionSpeech Enhancement

Best Practices for Noise-Based Augmentation to Improve the Performance of Deployable Speech-Based Emotion Recognition Systems

2021-04-18 · Mimansa Jaiswal, Emily Mower Provost

Speech emotion recognition is an important component of any human centered system. But speech characteristics produced and perceived by a person can be influenced by a multitude of reasons, both desirable such as emotion…

Adversarial AttackAutomatic Speech RecognitionData AugmentationEmotion Recognition+4

Investigations on Audiovisual Emotion Recognition in Noisy Conditions

2021-03-02 · Michael Neumann, Ngoc Thang Vu

In this paper we explore audiovisual emotion recognition under noisy acoustic conditions with a focus on speech features. We attempt to answer the following research questions: (i) How does speech emotion recognition per…

Emotion RecognitionSpeech Emotion Recognition

Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing

2026-01-26 · Marouane El Hizabri, Abdelfattah Bezzaz, Ismail Hayoukane, Youssef Taki arxiv

Speech Emotion Recognition systems often use static features like Mel-Frequency Cepstral Coefficients (MFCCs), Zero Crossing Rate (ZCR), and Root Mean Square Energy (RMSE). Because of this, they can misclassify emotions …

Speech Emotion RecognitionEmotion Classification