paper-with-me

홈 › Papers

Data Cleansing with Contrastive Learning for Vocal Note Event Annotations

2020-08-05 · Gabriel Meseguer-Brocal, Rachel Bittner, Simon Durand, Brian Brost

Data cleansing is a well studied strategy for cleaning erroneous labels in datasets, which has not yet been widely adopted in Music Information Retrieval. Previously proposed data cleansing models do not consider structured (e.g. time varying) labels, such as those common to music data. We propose a novel data cleansing model for time-varying, structured labels which exploits the local structure of the labels, and demonstrate its usefulness for vocal note event annotations in music. %Our model is trained in a contrastive learning manner by automatically creating local deformations of likely correct labels. Our model is trained in a contrastive learning manner by automatically contrasting likely correct labels pairs against local deformations of them. We demonstrate that the accuracy of a transcription model improves greatly when trained using our proposed strategy compared with the accuracy when trained using the original dataset. Additionally we use our model to estimate the annotation error rates in the DALI dataset, and highlight other potential uses for this type of model.

📄 PDF Abstract BibTeX arXiv:2008.02069

Code (1)

gabolsgabs/contrastive-data-cleansing 공식 구현

Tasks

Contrastive LearningInformation RetrievalMusic Information RetrievalRetrieval

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Metrical-accent Aware Vocal Onset Detection in Polyphonic Audio

2017-07-19 · Georgi Dzhambazov, Andre Holzapfel, Ajay Srinivasamurthy, Xavier Serra

The goal of this study is the automatic detection of onsets of the singing voice in polyphonic audio recordings. Starting with a hypothesis that the knowledge of the current position in a metrical cycle (i.e. metrical ac…

Onset DetectionPosition

Automatic recognition of element classes and boundaries in the birdsong with variable sequences

2016-01-23 · Takuya Koumura, Kazuo Okanoya

Researches on sequential vocalization often require analysis of vocalizations in long continuous sounds. In such studies as developmental ones or studies across generations in which days or months of vocalizations must b…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Boundary DetectionGeneral Classification+2

Phoneme-Informed Note Segmentation of Monophonic Vocal Music

2021-11-01 · NLP4MusA 2021 11 · Yukun Li, Emir Demirel, Polina Proutskova, Simon Dixon

Improving Query-by-Vocal Imitation with Contrastive Learning and Audio Pretraining

2024-08-21 · Jonathan Greif, Florian Schmid, Paul Primus, Gerhard Widmer

Query-by-Vocal Imitation (QBV) is about searching audio files within databases using vocal imitations created by the user's voice. Since most humans can effectively communicate sound concepts through voice, QBV offers th…

Contrastive Learning

Automatic Transcription of Flamenco Singing from Polyphonic Music Recordings

2015-10-14 · Kroher Nadine, Gómez Emilia

Automatic note-level transcription is considered one of the most challenging tasks in music information retrieval. The specific case of flamenco singing transcription poses a particular challenge due to its complex melod…

Information RetrievalMusic Information RetrievalOnset DetectionRetrieval