paper-with-me

Dusha

Dusha Crowd, Dusha Podcast

홈페이지 · 논문 2편

Dusha is a dataset for speech emotion recognition (SER) tasks. The corpus contains approximately 350 hours of data, more than 300 000 audio recordings with Russian speech and their transcripts. It is annotated using a crowd-sourcing platform and includes two subsets: acted and real-life. Source: Large Raw Emotional Dataset with Aggregation Mechanism

TextsAudio Russian

벤치마크

Speech Emotion Recognition on Dusha Crowd 결과 2개
Speech Emotion Recognition on Dusha Podcast 결과 2개