paper-with-me

홈 › Papers

Breaking Resource Barriers in Speech Emotion Recognition via Data Distillation

2024-06-21 · Yi Chang, Zhao Ren, Zhonghao Zhao, Thanh Tam Nguyen, Kun Qian, Tanja Schultz, Björn W. Schuller

Speech emotion recognition (SER) plays a crucial role in human-computer interaction. The emergence of edge devices in the Internet of Things (IoT) presents challenges in constructing intricate deep learning models due to constraints in memory and computational resources. Moreover, emotional speech data often contains private information, raising concerns about privacy leakage during the deployment of SER models. To address these challenges, we propose a data distillation framework to facilitate efficient development of SER models in IoT applications using a synthesised, smaller, and distilled dataset. Our experiments demonstrate that the distilled dataset can be effectively utilised to train SER models with fixed initialisation, achieving performances comparable to those developed using the original full emotional speech dataset.

📄 PDF Abstract BibTeX arXiv:2406.15119

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion Recognition

Similar Papers 제목 키워드 기반

Emotion Recognition in Low-Resource Settings: An Evaluation of Automatic Feature Selection Methods

2019-08-28 · Fasih Haider, Senja Pollak, Pierre Albert, Saturnino Luz

Research in automatic affect recognition has seldom addressed the issue of computational resource utilization. With the advent of ambient intelligence technology which employs a variety of low-power, resource-constrained…

Emotion Recognitionfeature selection

Divide-and-Conquer based Ensemble to Spot Emotions in Speech using MFCC and Random Forest

2016-10-05 · Abdul Malik Badshah, Jamil Ahmad, Mi Young Lee, Sung Wook Baik

Besides spoken words, speech signals also carry information about speaker gender, age, and emotional state which can be used in a variety of speech analysis applications. In this paper, a divide and conquer strategy for …

Emotion Recognition

OSUM: Advancing Open Speech Understanding Models with Limited Resources in Academia

2025-01-23 · Xuelong Geng, Kun Wei, Qijie Shao, Shuiyun Liu 외

Large Language Models (LLMs) have made significant progress in various downstream tasks, inspiring the development of Speech Understanding Language Models (SULMs) to enable comprehensive speech-based interactions. Howeve…

Emotion RecognitionEvent DetectionGender ClassificationSpeech Emotion Recognition+3

Breaking Down Power Barriers in On-Device Streaming ASR: Insights and Solutions

2024-02-20 · Yang Li, Yuan Shangguan, Yuhao Wang, Liangzhen Lai 외

Power consumption plays a crucial role in on-device streaming speech recognition, significantly influencing the user experience. This study explores how the configuration of weight parameters in speech recognition models…

speech-recognitionSpeech Recognition

Light-SERNet: A lightweight fully convolutional neural network for speech emotion recognition

2021-10-07 · Arya Aftab, Alireza Morsali, Shahrokh Ghaemmaghami, Benoit Champagne

Detecting emotions directly from a speech signal plays an important role in effective human-computer interactions. Existing speech emotion recognition models require massive computational and storage resources, making th…

Emotion RecognitionSpeech Emotion Recognition