paper-with-me

홈 › Papers

Knowledge-based Multimodal Music Similarity

2023-06-21 · Andrea Poltronieri

Music similarity is an essential aspect of music retrieval, recommendation systems, and music analysis. Moreover, similarity is of vital interest for music experts, as it allows studying analogies and influences among composers and historical periods. Current approaches to musical similarity rely mainly on symbolic content, which can be expensive to produce and is not always readily available. Conversely, approaches using audio signals typically fail to provide any insight about the reasons behind the observed similarity. This research addresses the limitations of current approaches by focusing on the study of musical similarity using both symbolic and audio content. The aim of this research is to develop a fully explainable and interpretable system that can provide end-users with more control and understanding of music similarity and classification systems.

📄 PDF Abstract BibTeX arXiv:2306.12249

Code (0)

등록된 구현이 없습니다.

Tasks

Recommendation SystemsRetrieval

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

Multimodal Lyrics-Rhythm Matching

2023-01-06 · Callie C. Liao, Duoduo Liao, Jesse Guessford

Despite the recent increase in research on artificial intelligence for music, prominent correlations between key components of lyrics and rhythm such as keywords, stressed syllables, and strong beats are not frequently s…

Rhythm

MMVA: Multimodal Matching Based on Valence and Arousal across Images, Music, and Musical Captions

2025-01-02 · Suhwan Choi, Kyu Won Kim, Myungjoo Kang

We introduce Multimodal Matching based on Valence and Arousal (MMVA), a tri-modal encoder framework designed to capture emotional content across images, music, and musical captions. To support this framework, we expand t…

Evaluation of pretrained language models on music understanding

2024-09-17 · Yannis Vasilakis, Rachel Bittner, Johan Pauwels

Music-text multimodal systems have enabled new approaches to Music Information Research (MIR) applications such as audio-to-text and text-to-audio retrieval, text-based song generation, and music captioning. Despite the …

Music CaptioningNegationSensitivityText to Audio Retrieval+1

Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model

2023-11-02 · Jaeyong Kang, Soujanya Poria, Dorien Herremans

Numerous studies in the field of music generation have demonstrated impressive performance, yet virtually no models are able to directly generate music to match accompanying videos. In this work, we develop a generative …

Music GenerationRhythm

Music's Multimodal Complexity in AVQA: Why We Need More than General Multimodal LLMs

2025-05-27 · Wenhao You, Xingjian Diao, Chunhui Zhang, Keyi Kong 외

While recent Multimodal Large Language Models exhibit impressive capabilities for general multimodal tasks, specialized domains like music necessitate tailored approaches. Music Audio-Visual Question Answering (Music AVQ…

Audio-visual Question AnsweringQuestion AnsweringVisual Question Answering