paper-with-me

Papers

Self-Supervised Contrastive Learning for Robust Audio-Sheet Music Retrieval Systems

2023-09-21 · Luis Carvalho, Tobias Washüttl, Gerhard Widmer

Linking sheet music images to audio recordings remains a key problem for the development of efficient cross-modal music retrieval systems. One of the fundamental approaches toward this task is to learn a cross-modal embedding space via deep neural networks that is able to connect short snippets of audio and sheet music. However, the scarcity of annotated data from real musical content affects the capability of such methods to generalize to real retrieval scenarios. In this work, we investigate whether we can mitigate this limitation with self-supervised contrastive learning, by exposing a network to a large amount of real music data as a pre-training step, by contrasting randomly augmented views of snippets of both modalities, namely audio and sheet images. Through a number of experiments on synthetic and real piano data, we show that pre-trained models are able to retrieve snippets with better precision in all scenarios and pre-training configurations. Encouraged by these results, we employ the snippet embeddings in the higher-level task of cross-modal piece identification and conduct more experiments on several retrieval configurations. In this task, we observe that the retrieval quality improves from 30% up to 100% when real music data is present. We then conclude by arguing for the potential of self-supervised contrastive learning for alleviating the annotated data scarcity in multi-modal music retrieval models.

📄 PDF Abstract BibTeX arXiv:2309.12134

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningRetrieval

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Unsupervised Generative Adversarial Alignment Representation for Sheet music, Audio and Lyrics

2020-07-29 · Donghuo Zeng, Yi Yu, Keizo Oyama

Sheet music, audio, and lyrics are three main modalities during writing a song. In this paper, we propose an unsupervised generative adversarial alignment representation (UGAAR) model to learn deep discriminative represe…

Representation Learning

Towards Proper Contrastive Self-supervised Learning Strategies For Music Audio Representation

2022-07-10 · Jeong Choi, Seongwon Jang, Hyunsouk Cho, Sehee Chung

The common research goal of self-supervised learning is to extract a general representation which an arbitrary downstream task would benefit from. In this work, we investigate music audio representation learned from diff…

Contrastive LearningInformation RetrievalMusic Information RetrievalRetrieval+1

Learning Audio - Sheet Music Correspondences for Score Identification and Offline Alignment

2017-07-31 · Dorfer Matthias, Arzt Andreas, Widmer Gerhard

This work addresses the problem of matching short excerpts of audio with their respective counterparts in sheet music images. We show how to employ neural network-based cross-modality embedding spaces for solving the fol…

Passage Summarization with Recurrent Models for Audio-Sheet Music Retrieval

2023-09-21 · Luis Carvalho, Gerhard Widmer

Many applications of cross-modal music retrieval are related to connecting sheet music images to audio recordings. A typical and recent approach to this is to learn, via deep neural networks, a joint embedding space that…

Retrieval

Towards Score Following in Sheet Music Images

2016-12-15 · Matthias Dorfer, Andreas Arzt, Gerhard Widmer

This paper addresses the matching of short music audio snippets to the corresponding pixel location in images of sheet music. A system is presented that simultaneously learns to read notes, listens to music and matches t…

Position