paper-with-me

홈 › Papers

SEWA DB: A Rich Database for Audio-Visual Emotion and Sentiment Research in the Wild

2019-01-09 · Jean Kossaifi, Robert Walecki, Yannis Panagakis, Jie Shen, Maximilian Schmitt, Fabien Ringeval, Jing Han, Vedhas Pandit, Antoine Toisoul, Bjorn Schuller, Kam Star, Elnar Hajiyev, Maja Pantic

Natural human-computer interaction and audio-visual human behaviour sensing systems, which would achieve robust performance in-the-wild are more needed than ever as digital devices are increasingly becoming an indispensable part of our life. Accurately annotated real-world data are the crux in devising such systems. However, existing databases usually consider controlled settings, low demographic variability, and a single task. In this paper, we introduce the SEWA database of more than 2000 minutes of audio-visual data of 398 people coming from six cultures, 50% female, and uniformly spanning the age range of 18 to 65 years old. Subjects were recorded in two different contexts: while watching adverts and while discussing adverts in a video chat. The database includes rich annotations of the recordings in terms of facial landmarks, facial action units (FAU), various vocalisations, mirroring, and continuously valued valence, arousal, liking, agreement, and prototypic examples of (dis)liking. This database aims to be an extremely valuable resource for researchers in affective computing and automatic human sensing and is expected to push forward the research in human behaviour analysis, including cultural studies. Along with the database, we provide extensive baseline experiments for automatic FAU detection and automatic valence, arousal and (dis)liking intensity estimation.

📄 PDF Abstract BibTeX arXiv:1901.02839

Code (1)

pwc-1/Paper-10/tree/main/sew_d mindspore

Similar Papers 제목 키워드 기반

Construction of Japanese Audio-Visual Emotion Database and Its Application in Emotion Recognition

2016-05-01 · LREC 2016 5 · Nurul Lubis, R Gomez, y, Sakriani Sakti 외

Emotional aspects play a vital role in making human communication a rich and dynamic experience. As we introduce more automated system in our daily lives, it becomes increasingly important to incorporate emotion to provi…

Emotion Recognition

On the use of Self-supervised Pre-trained Acoustic and Linguistic Features for Continuous Speech Emotion Recognition

2020-11-18 · Manon Macary, Marie Tahon, Yannick Estève, Anthony Rousseau

Pre-training for feature extraction is an increasingly studied approach to get better continuous representations of audio and text content. In the present work, we use wav2vec and camemBERT as self-supervised learned mod…

Emotion RecognitionSpeech Emotion Recognition

Naturalistic Audio-Visual Emotion Database

2014-12-01 · WS 2014 12 · Sudarsana Reddy Kadiri, P. Gangamohan, V. K. Mittal, B. Yegnanarayana
Emotion Recognition

Enriching Multimodal Sentiment Analysis through Textual Emotional Descriptions of Visual-Audio Content

2024-12-12 · Sheng Wu, Xiaobao Wang, Longbiao Wang, Dongxiao He 외

Multimodal Sentiment Analysis (MSA) stands as a critical research frontier, seeking to comprehensively unravel human emotions by amalgamating text, audio, and visual data. Yet, discerning subtle emotional nuances within …

Multimodal Sentiment AnalysisSentiment Analysis

Measuring Mother-Infant Emotions By Audio Sensing

2019-12-10 · Xuewen Yao, Dong He, Tiancheng Jing, Kaya de Barbaro

It has been suggested in developmental psychology literature that the communication of affect between mothers and their infants correlates with the socioemotional and cognitive development of infants. In this study, we o…

Active Learning