paper-with-me

Papers

Source Separation of Small Classical Ensembles: Challenges and Opportunities

2025-05-23 · Gerardo Roa-Dabike, Trevor J. Cox, Jon P. Barker, Michael A. Akeroyd, Scott Bannister, Bruno Fazenda, Jennifer Firth, Simone Graetzer, Alinka Greasley, Rebecca R. Vos, William M. Whitmer

Musical (MSS) source separation of western popular music using non-causal deep learning can be very effective. In contrast, MSS for classical music is an unsolved problem. Classical ensembles are harder to separate than popular music because of issues such as the inherent greater variation in the music; the sparsity of recordings with ground truth for supervised training; and greater ambiguity between instruments. The Cadenza project has been exploring MSS for classical music. This is being done so music can be remixed to improve listening experiences for people with hearing loss. To enable the work, a new database of synthesized woodwind ensembles was created to overcome instrumental imbalances in the EnsembleSet. For the MSS, a set of ConvTasNet models was used with each model being trained to extract a string or woodwind instrument. ConvTasNet was chosen because it enabled both causal and non-causal approaches to be tested. Non-causal approaches have dominated MSS work and are useful for recorded music, but for live music or processing on hearing aids, causal signal processing is needed. The MSS performance was evaluated on the two small datasets (Bach10 and URMP) of real instrument recordings where the ground-truth is available. The performances of the causal and non-causal systems were similar. Comparing the average Signal-to-Distortion (SDR) of the synthesized validation set (6.2 dB causal; 6.9 non-causal), to the real recorded evaluation set (0.3 dB causal, 0.4 dB non-causal), shows that mismatch between synthesized and recorded data is a problem. Future work needs to either gather more real recordings that can be used for training, or to improve the realism and diversity of the synthesized recordings to reduce the mismatch...

📄 PDF Abstract BibTeX arXiv:2505.17823

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ConvTasNet 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Deep Learning Based Source Separation Applied To Choir Ensembles

2020-08-17 · Darius Petermann, Pritish Chandna, Helena Cuesta, Jordi Bonada 외

Choral singing is a widely practiced form of ensemble singing wherein a group of people sing simultaneously in polyphonic harmony. The most commonly practiced setting for choir ensembles consists of four parts; Soprano, …

Deep Learning

The Spheres Dataset: Multitrack Orchestral Recordings for Music Source Separation and Information Retrieval

2025-11-26 · Jaime Garcia-Martinez, David Diaz-Guerra, John Anderson, Ricardo Falcon-Perez 외 arxiv

This paper introduces The Spheres dataset, multitrack orchestral recordings designed to advance machine learning research in music source separation and related MIR tasks within the classical music domain. The dataset is…

Music Source SeparationInformation Retrieval

Quantum-Classical Separations in Shallow-Circuit-Based Learning with and without Noises

2024-05-01 · Zhihan Zhang, Weiyuan Gong, Weikang Li, Dong-Ling Deng

We study quantum-classical separations between classical and quantum supervised learning models based on constant depth (i.e., shallow) circuits, in scenarios with and without noises. We construct a classification proble…

Exploiting Temporal Structures of Cyclostationary Signals for Data-Driven Single-Channel Source Separation

2022-08-22 · Gary C. F. Lee, Amir Weiss, Alejandro Lancho, Jennifer Tang 외

We study the problem of single-channel source separation (SCSS), and focus on cyclostationary signals, which are particularly suitable in a variety of application domains. Unlike classical SCSS approaches, we consider a …

The Chamber Ensemble Generator: Limitless High-Quality MIR Data via Generative Modeling

2022-09-28 · Yusong Wu, Josh Gardner, Ethan Manilow, Ian Simon 외

Data is the lifeblood of modern machine learning systems, including for those in Music Information Retrieval (MIR). However, MIR has long been mired by small datasets and unreliable labels. In this work, we propose to br…

Information RetrievalMulti-instrument Music TranscriptionMusic Information RetrievalMusic Transcription+1