paper-with-me

홈 › Papers

Transcription Is All You Need: Learning to Separate Musical Mixtures with Score as Supervision

2020-10-22 · Yun-Ning Hung, Gordon Wichern, Jonathan Le Roux

Most music source separation systems require large collections of isolated sources for training, which can be difficult to obtain. In this work, we use musical scores, which are comparatively easy to obtain, as a weak label for training a source separation system. In contrast with previous score-informed separation approaches, our system does not require isolated sources, and score is used only as a training target, not required for inference. Our model consists of a separator that outputs a time-frequency mask for each instrument, and a transcriptor that acts as a critic, providing both temporal and frequency supervision to guide the learning of the separator. A harmonic mask constraint is introduced as another way of leveraging score information during training, and we propose two novel adversarial losses for additional fine-tuning of both the transcriptor and the separator. Results demonstrate that using score information outperforms temporal weak-labels, and adversarial structures lead to further improvements in both separation and transcription performance.

📄 PDF Abstract BibTeX arXiv:2010.11904

Code (0)

등록된 구현이 없습니다.

Tasks

AllMusic Source Separation

Similar Papers 제목 키워드 기반

Simultaneous Separation and Transcription of Mixtures with Multiple Polyphonic and Percussive Instruments

2019-10-22 · Ethan Manilow, Prem Seetharaman, Bryan Pardo

We present a single deep learning architecture that can both separate an audio recording of a musical mixture into constituent single-instrument recordings and transcribe these instruments into a human-readable format at…

Bespoke Neural Networks for Score-Informed Source Separation

2020-09-29 · Ethan Manilow, Bryan Pardo

In this paper, we introduce a simple method that can separate arbitrary musical instruments from an audio mixture. Given an unaligned MIDI transcription for a target instrument from an input mixture, we synthesize new mi…

Musical Rhythm Transcription Based on Bayesian Piece-Specific Score Models Capturing Repetitions

2019-08-18 · Eita Nakamura, Kazuyoshi Yoshii

Most work on musical score models (a.k.a. musical language models) for music transcription has focused on describing the local sequential dependence of notes in musical scores and failed to capture their global repetitiv…

Computational EfficiencyLanguage ModellingMusic TranscriptionRhythm

Evaluating Non-aligned Musical Score Transcriptions with MV2H

2019-06-03 · Andrew McLeod

The original MV2H metric was designed to evaluate systems which transcribe from an input audio (or MIDI) piece to a complete musical score. However, it requires both the transcribed score and the ground truth score to be…

Musical Source Separation of Brazilian Percussion

2025-03-06 · Richa Namballa, Giovana Morais, Magdalena Fuentes

Musical source separation (MSS) has recently seen a big breakthrough in separating instruments from a mixture in the context of Western music, but research on non-Western instruments is still limited due to a lack of dat…