paper-with-me

Papers

Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings

2024-01-26 · Shreyan Chowdhury, Gerhard Widmer

This paper explores a specific sub-task of cross-modal music retrieval. We consider the delicate task of retrieving a performance or rendition of a musical piece based on a description of its style, expressive character, or emotion from a set of different performances of the same piece. We observe that a general purpose cross-modal system trained to learn a common text-audio embedding space does not yield optimal results for this task. By introducing two changes -- one each to the text encoder and the audio encoder -- we demonstrate improved performance on a dataset of piano performances and associated free-text descriptions. On the text side, we use emotion-enriched word embeddings (EWE) and on the audio side, we extract mid-level perceptual features instead of generic audio embeddings. Our results highlight the effectiveness of mid-level perceptual features learnt from music and emotion enriched word embeddings learnt from emotion-labelled text in capturing musical expression in a cross-modal setting. Additionally, our interpretable mid-level features provide a route for introducing explainability in the retrieval and downstream recommendation processes.

📄 PDF Abstract BibTeX arXiv:2401.14826

Code (0)

등록된 구현이 없습니다.

Tasks

RetrievalWord Embeddings

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Towards Explaining Expressive Qualities in Piano Recordings: Transfer of Explanatory Features via Acoustic Domain Adaptation

2021-02-26 · Shreyan Chowdhury, Gerhard Widmer

Emotion and expressivity in music have been topics of considerable interest in the field of music information retrieval. In recent years, mid-level perceptual features have been suggested as means to explain computationa…

DiversityDomain AdaptationInformation RetrievalMusic Information Retrieval+2

Expressive Music Data Processing and Generation

2025-03-14 · Jingwei Liu

Musical expressivity and coherence are indispensable in music composition and performance, while often neglected in modern AI generative models. In this work, we introduce a listening-based data-processing technique that…

Exploiting Device and Audio Data to Tag Music with User-Aware Listening Contexts

2022-11-14 · Karim M. Ibrahim, Elena V. Epure, Geoffroy Peeters, Gaël Richard

As music has become more available especially on music streaming platforms, people have started to have distinct preferences to fit to their varying listening situations, also known as context. Hence, there has been a gr…

RetrievalTAG

Toward Universal Text-to-Music Retrieval

2022-11-26 · Seungheon Doh, Minz Won, Keunwoo Choi, Juhan Nam

This paper introduces effective design choices for text-to-music retrieval systems. An ideal text-based retrieval system would support various input queries such as pre-defined tags, unseen tags, and sentence-level descr…

Music ClassificationRetrievalSentenceTAG

BeatDance: A Beat-Based Model-Agnostic Contrastive Learning Framework for Music-Dance Retrieval

2023-10-16 · Kaixing Yang, Xukun Zhou, Xulong Tang, Ran Diao 외

Dance and music are closely related forms of expression, with mutual retrieval between dance videos and music being a fundamental task in various fields like education, art, and sports. However, existing methods often su…

Contrastive LearningRetrieval