paper-with-me

Papers

Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning

2024-12-26 · Dengming Zhang, Weitao You, Ziheng Liu, Lingyun Sun, Pei Chen

Dynamic Music Emotion Recognition (DMER) aims to predict the emotion of different moments in music, playing a crucial role in music information retrieval. The existing DMER methods struggle to capture long-term dependencies when dealing with sequence data, which limits their performance. Furthermore, these methods often overlook the influence of individual differences on emotion perception, even though everyone has their own personalized emotional perception in the real world. Motivated by these issues, we explore more effective sequence processing methods and introduce the Personalized DMER (PDMER) problem, which requires models to predict emotions that align with personalized perception. Specifically, we propose a Dual-Scale Attention-Based Meta-Learning (DSAML) method. This method fuses features from a dual-scale feature extractor and captures both short and long-term dependencies using a dual-scale attention transformer, improving the performance in traditional DMER. To achieve PDMER, we design a novel task construction strategy that divides tasks by annotators. Samples in a task are annotated by the same annotator, ensuring consistent perception. Leveraging this strategy alongside meta-learning, DSAML can predict personalized perception of emotions with just one personalized annotation sample. Our objective and subjective experiments demonstrate that our method can achieve state-of-the-art performance in both traditional DMER and PDMER.

📄 PDF Abstract BibTeX arXiv:2412.19200

Code (1)

Littleor/Personalized-DMER 공식 구현 pytorch

Tasks

Emotion RecognitionInformation RetrievalMeta-LearningMusic Emotion RecognitionMusic Information Retrieval

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

EmoHeal: An End-to-End System for Personalized Therapeutic Music Retrieval from Fine-grained Emotions

2025-09-19 · Xinchen Wan, Jinhua Liang, Huan Zhang arxiv

Existing digital mental wellness tools often overlook the nuanced emotional states underlying everyday challenges. For example, pre-sleep anxiety affects more than 1.5 billion people worldwide, yet current approaches rem…

Emotion Recognition

A Efficient Multimodal Framework for Large Scale Emotion Recognition by Fusing Music and Electrodermal Activity Signals

2020-08-22 · Guanghao Yin, Shou-qian Sun, Dian Yu, Dejian Li 외

Considerable attention has been paid for physiological signal-based emotion recognition in field of affective computing. For the reliability and user friendly acquisition, Electrodermal Activity (EDA) has great advantage…

Emotion Recognition

Emotion-Aware Music Recommendation System: Enhancing User Experience Through Real-Time Emotional Context

2023-11-17 · Tina Babu, Rekha R Nair, Geetha A

This study addresses the deficiency in conventional music recommendation systems by focusing on the vital role of emotions in shaping users music choices. These systems often disregard the emotional context, relying pred…

Music RecommendationRecommendation Systems

Affective Music Information Retrieval

2015-02-18 · Wang Ju-Chiang, Yang Yi-Hsuan, Wang Hsin-Min

Much of the appeal of music lies in its power to convey emotions/moods and to evoke them in listeners. In consequence, the past decade witnessed a growing interest in modeling emotions from musical signals in the music i…

Emotion RecognitionInformation RetrievalMusic Information RetrievalRetrieval

Secure & Personalized Music-to-Video Generation via CHARCHA

2025-02-03 · Mehul Agarwal, Gauri Agarwal, Santiago Benoit, Andrew Lippman 외

Music is a deeply personal experience and our aim is to enhance this with a fully-automated pipeline for personalized music video generation. Our work allows listeners to not just be consumers but co-creators in the musi…

RhythmVideo Generation