paper-with-me

홈 › Papers

Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature

2013-03-07 · Mayank Bhargava, Tim Polzehl

This paper is devoted to improve automatic emotion recognition from speech by incorporating rhythm and temporal features. Research on automatic emotion recognition so far has mostly been based on applying features like MFCCs, pitch and energy or intensity. The idea focuses on borrowing rhythm features from linguistic and phonetic analysis and applying them to the speech signal on the basis of acoustic knowledge only. In addition to this we exploit a set of temporal and loudness features. A segmentation unit is employed in starting to separate the voiced/unvoiced and silence parts and features are explored on different segments. Thereafter different classifiers are used for classification. After selecting the top features using an IGR filter we are able to achieve a recognition rate of 80.60 % on the Berlin Emotion Database for the speaker dependent framework.

📄 PDF Abstract BibTeX arXiv:1303.1761

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionRhythm

Similar Papers 제목 키워드 기반

DARS: Dysarthria-Aware Rhythm-Style Synthesis for ASR Enhancement

2026-03-02 · Minghui Wu, Xueling Liu, Jiahuan Fan, Haitao Tang 외 arxiv

Dysarthric speech exhibits abnormal prosody and significant speaker variability, presenting persistent challenges for automatic speech recognition (ASR). While text-to-speech (TTS)-based data augmentation has shown poten…

Speech RecognitionData Augmentation

EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation

2023-05-30 · Xingqun Qi, Chen Liu, Lincheng Li, Jie Hou 외

Generating vivid and diverse 3D co-speech gestures is crucial for various applications in animating virtual avatars. While most existing methods can generate gestures from audio directly, they usually overlook that emoti…

Gesture GenerationRhythm

Imperceptible Rhythm Backdoor Attacks: Exploring Rhythm Transformation for Embedding Undetectable Vulnerabilities on Speech Recognition

2024-06-16 · Wenhan Yao, Jiangkun Yang, Yongqiang He, Jia Liu 외

Speech recognition is an essential start ring of human-computer interaction, and recently, deep learning models have achieved excellent success in this task. However, when the model training and private data provider are…

Automatic Speech RecognitionData PoisoningRhythmSpeaker Verification+2

Rhythm Features for Speaker Identification

2025-06-07 · Nick Mehlman, Thomas Thebaud, Dani Byrd, Shri Narayanan

While deep learning models have demonstrated robust performance in speaker recognition tasks, they primarily rely on low-level audio features learned empirically from spectrograms or raw waveforms. However, prior work ha…

Deep LearningRhythmSpeaker IdentificationSpeaker Recognition

Speech-Emotion Detection in an Indonesian Movie

2020-05-01 · LREC 2020 5 · Fahmi Fahmi, Meganingrum Arista Jiwanggi, Mirna Adriani

The growing demand to develop an automatic emotion recognition system for the Human-Computer Interaction field had pushed some research in speech emotion detection. Although it is growing, there is still little research …

Emotion RecognitionGeneral Classification