paper-with-me

Papers

Hybrid Spectrogram and Waveform Source Separation

2021-11-05 · Alexandre Défossez

Source separation models either work on the spectrogram or waveform domain. In this work, we show how to perform end-to-end hybrid source separation, letting the model decide which domain is best suited for each source, and even combining both. The proposed hybrid version of the Demucs architecture won the Music Demixing Challenge 2021 organized by Sony. This architecture also comes with additional improvements, such as compressed residual branches, local attention or singular value regularization. Overall, a 1.4 dB improvement of the Signal-To-Distortion (SDR) was observed across all sources as measured on the MusDB HQ dataset, an improvement confirmed by human subjective evaluation, with an overall quality rated at 2.83 out of 5 (2.36 for the non hybrid Demucs), and absence of contamination at 3.04 (against 2.37 for the non hybrid Demucs and 2.44 for the second ranking model submitted at the competition).

📄 PDF Abstract BibTeX arXiv:2111.03600

Code (1)

facebookresearch/demucs 공식 구현 pytorch

Tasks

Music Source Separation

Similar Papers 제목 키워드 기반

Hybrid Y-Net Architecture for Singing Voice Separation

2023-03-05 · Rashen Fernando, Pamudu Ranasinghe, Udula Ranasinghe, Janaka Wijayakulasooriya 외

This research paper presents a novel deep learning-based neural network architecture, named Y-Net, for achieving music source separation. The proposed architecture performs end-to-end hybrid source separation by extracti…

Music Source Separation

End-to-end music source separation: is it possible in the waveform domain?

2018-10-29 · Francesc Lluís, Jordi Pons, Xavier Serra

Most of the currently successful source separation techniques use the magnitude spectrogram as input, and are therefore by default omitting part of the signal: the phase. To avoid omitting potentially useful information,…

Deep LearningMusic Source Separation

Real-time Low-latency Music Source Separation using Hybrid Spectrogram-TasNet

2024-02-27 · Satvik Venkatesh, Arthur Benilov, Philip Coleman, Frederic Roskam

There have been significant advances in deep learning for music demixing in recent years. However, there has been little attention given to how these neural networks can be adapted for real-time low-latency applications,…

Music Source Separation

End-to-end Networks for Supervised Single-channel Speech Separation

2018-10-05 · Shrikant Venkataramani, Paris Smaragdis

The performance of single channel source separation algorithms has improved greatly in recent times with the development and deployment of neural networks. However, many such networks continue to operate on the magnitude…

Speech Separation

Danna-Sep: Unite to separate them all

2021-12-07 · Chin-Yun Yu, Kin-Wai Cheuk

Deep learning-based music source separation has gained a lot of interest in the last decades. Most of the existing methods operate with either spectrograms or waveforms. Spectrogram based models learn suitable masks for …

AllMusic Source Separation