paper-with-me

홈 › Papers

Speech Emotion Recognition using Attention-based LSTM-Network with Residual Connection

2026-06-02 · Daniil Krasnoproshin, Maxim Vashkevich arxiv

Speech emotion recognition is an important component of modern human-computer interaction systems. However, many state-of-the-art approaches rely on large pretrained models with high computational and memory requirements, limiting their applicability. This paper proposes ResLSTM-SA, a lightweight architecture that integrates residual connections with soft attention within an LSTM-based framework. Evaluated on the RAVDESS dataset under strict speaker-independent partitioning, the proposed model outperforms conventional attention-based LSTM baselines and several previously reported CNN- and hybrid CNN-LSTM architectures in terms of unweighted average recall (UAR). The best-performing variant (ResLSTM-SA-h64) achieves a maximum UAR of 0.6517 with only 46.8k trainable parameters, delivering competitive accuracy with three orders of magnitude fewer parameters than large-scale self-supervised alternatives, thereby enabling efficient deployment on edge devices and real-time voice assistants. The source code is available at https://github.com/Mak-Sim/ResLSTM-SER.

📄 PDF Abstract BibTeX arXiv:2606.03359

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Emotion Recognition

Similar Papers 제목 키워드 기반

Multi-stream Attention-based BLSTM with Feature Segmentation for Speech Emotion Recognition

2020-10-25 · Interspeech 2020 10 · Yuya Chiba1, Takashi Nose1, Akinori Ito

This paper proposes a speech emotion recognition technique that considers the suprasegmental characteristics and temporal change of individual speech parameters. In recent years, speech emotion recognition using Bidir…

Data AugmentationEmotional Speech SynthesisEmotion RecognitionSpeech Emotion Recognition+1

Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion

2026-09-04 · Xu Lin, Ke Wang, Hui Kang, Xinying Wang arxiv

Multimodal emotion recognition has attracted growing interest due to its importance in human-computer interaction, remote education, and healthcare. This paper proposes a novel multimodal emotion recognition framework th…

Multimodal Emotion Recognition

Efficient Arabic emotion recognition using deep neural networks

2020-10-31 · Ahmed Ali, Yasser Hifny

Emotion recognition from speech signal based on deep learning is an active research area. Convolutional neural networks (CNNs) may be the dominant method in this area. In this paper, we implement two neural architectures…

Emotion RecognitionSpeech Emotion Recognition

Speech Emotion Recognition Based on CNN+LSTM Model

2021-10-01 · ROCLING 2021 10 · Wei Mou, Pei-Hsuan Shen, Chu-Yun Chu, Yu-Cheng Chiu 외

Due to the popularity of intelligent dialogue assistant services, speech emotion recognition has become more and more important. In the communication between humans and machines, emotion recognition and emotion analysis …

Emotion RecognitionmodelSpeech Emotion Recognition

Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model

2026-04-16 · Adelekun Oluwademilade, Ademola Adedamola, Abiola Abdulhakeem, Akinpelu Azeezat 외 arxiv

Speech Emotion Recognition (SER) is the use of machines to detect the emotional state of humans based on the speech, which is gaining importance in natural human-computer interaction. Speech is a very valuable source of …

Speech Emotion Recognition