paper-with-me

Papers

SER_AMPEL: a multi-source dataset for speech emotion recognition of Italian older adults

2023-11-24 · Alessandra Grossi, Francesca Gasparini

In this paper, SER_AMPEL, a multi-source dataset for speech emotion recognition (SER) is presented. The peculiarity of the dataset is that it is collected with the aim of providing a reference for speech emotion recognition in case of Italian older adults. The dataset is collected following different protocols, in particular considering acted conversations, extracted from movies and TV series, and recording natural conversations where the emotions are elicited by proper questions. The evidence of the need for such a dataset emerges from the analysis of the state of the art. Preliminary considerations on the critical issues of SER are reported analyzing the classification results on a subset of the proposed dataset.

📄 PDF Abstract BibTeX arXiv:2311.14483

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion Recognition

Similar Papers 제목 키워드 기반

Learning Alignment for Multimodal Emotion Recognition from Speech

2019-09-06 · Haiyang Xu, HUI ZHANG, Kun Han, Yun Wang 외

Speech emotion recognition is a challenging problem because human convey emotions in subtle and complex ways. For emotion recognition on human speech, one can either extract emotion related features from audio signals or…

Emotion RecognitionMultimodal Emotion RecognitionSpeech Emotion Recognitionspeech-recognition+1

EmoHopeSpeech: An Annotated Dataset of Emotions and Hope Speech in English and Arabic

2025-05-17 · Wajdi Zaghouani, Md. Rafiul Biswas

This research introduces a bilingual dataset comprising 23,456 entries for Arabic and 10,036 entries for English, annotated for emotions and hope speech, addressing the scarcity of multi-emotion (Emotion and hope) datase…

Speech Emotion Diarization: Which Emotion Appears When?

2023-06-22 · Yingzhi Wang, Mirco Ravanelli, Alya Yacoubi

Speech Emotion Recognition (SER) typically relies on utterance-level solutions. However, emotions conveyed through speech should be considered as discrete speech events with definite temporal boundaries, rather than attr…

Emotion Recognitionspeaker-diarizationSpeaker DiarizationSpeech Emotion Recognition

EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses

2026-04-29 · Shuhao Xu, Yifan Hu, Jingjing Wu, Zhihao Du 외 arxiv

Emotion perception and adaptive expression are fundamental capabilities in human-agent interaction. While recent advances in speech emotion captioning (SEC) have improved fine-grained emotional modeling, existing systems…

Speech Synthesis

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition

2025-01-06 · Ruoyu Zhao, Xiantao Jiang, F. Richard Yu, Victor C. M. Leung 외

Speech Emotion Recognition (SER) plays a crucial role in enhancing human-computer interaction. Cross-Linguistic SER (CLSER) has been a challenging research problem due to significant variability in linguistic and acousti…

Emotion RecognitionSpeech Emotion RecognitionTransfer Learning