paper-with-me

Papers

Annotation-free Automatic Music Transcription with Scalable Synthetic Data and Adversarial Domain Confusion

2023-12-16 · Gakusei Sato, Taketo Akama

Automatic Music Transcription (AMT) is a vital technology in the field of music information processing. Despite recent enhancements in performance due to machine learning techniques, current methods typically attain high accuracy in domains where abundant annotated data is available. Addressing domains with low or no resources continues to be an unresolved challenge. To tackle this issue, we propose a transcription model that does not require any MIDI-audio paired data through the utilization of scalable synthetic audio for pre-training and adversarial domain confusion using unannotated real audio. In experiments, we evaluate methods under the real-world application scenario where training datasets do not include the MIDI annotation of audio in the target data domain. Our proposed method achieved competitive performance relative to established baseline methods, despite not utilizing any real datasets of paired MIDI-audio. Additionally, ablation studies have provided insights into the scalability of this approach and the forthcoming challenges in the field of AMT research.

📄 PDF Abstract BibTeX arXiv:2312.10402

Code (0)

등록된 구현이 없습니다.

Tasks

Music Transcription

Similar Papers 제목 키워드 기반

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models

2026-05-06 · Yukun Chen, Tianrui Wang, Zhaoxi Mu, Xinyu Yang 외 arxiv

High-quality singing annotations are fundamental to modern Singing Voice Synthesis (SVS) systems. However, obtaining these annotations at scale through manual labeling is unrealistic due to the substantial labor and musi…

Source Separation & Automatic Transcription for Music

2024-12-09 · Bradford Derby, Lucas Dunker, Samarth Galchar, Shashank Jarmale 외

Source separation is the process of isolating individual sounds in an auditory mixture of multiple sounds [1], and has a variety of applications ranging from speech enhancement and lyric transcription [2] to digital audi…

Music TranscriptionSpeech Enhancement

Count The Notes: Histogram-Based Supervision for Automatic Music Transcription

2025-11-18 · Jonathan Yaffe, Ben Maman, Meinard Müller, Amit H. Bermano arxiv

Automatic Music Transcription (AMT) converts audio recordings into symbolic musical representations. Training deep neural networks (DNNs) for AMT typically requires strongly aligned training pairs with precise frame-leve…

Music Transcription

Multi-Channel Automatic Music Transcription Using Tensor Algebra

2021-07-23 · Marmoret Axel, Bertin Nancy, Cohen Jeremy

Music is an art, perceived in unique ways by every listener, coming from acoustic signals. In the meantime, standards as musical scores exist to describe it. Even if humans can make this transcription, it is costly in te…

Music Transcriptiontensor algebra

Scorpiano -- A System for Automatic Music Transcription for Monophonic Piano Music

2021-08-24 · Bojan Sofronievski, Branislav Gerazov

Music transcription is the process of transcribing music audio into music notation. It is a field in which the machines still cannot beat human performance. The main motivation for automatic music transcription is to mak…

Music TranscriptionOnset Detection