paper-with-me

홈 › Papers

Music Mixing Style Transfer: A Contrastive Learning Approach to Disentangle Audio Effects

2022-11-04 · Junghyun Koo, Marco A. Martínez-Ramírez, Wei-Hsiang Liao, Stefan Uhlich, Kyogu Lee, Yuki Mitsufuji

We propose an end-to-end music mixing style transfer system that converts the mixing style of an input multitrack to that of a reference song. This is achieved with an encoder pre-trained with a contrastive objective to extract only audio effects related information from a reference music recording. All our models are trained in a self-supervised manner from an already-processed wet multitrack dataset with an effective data preprocessing method that alleviates the data scarcity of obtaining unprocessed dry data. We analyze the proposed encoder for the disentanglement capability of audio effects and also validate its performance for mixing style transfer through both objective and subjective evaluations. From the results, we show the proposed system not only converts the mixing style of multitrack audio close to a reference but is also robust with mixture-wise style transfer upon using a music source separation model.

📄 PDF Abstract BibTeX arXiv:2211.02247

Code (1)

jhtonykoo/music_mixing_style_transfer 공식 구현 pytorch

Tasks

Contrastive LearningDisentanglementMusic Source SeparationStyle Transfer

Similar Papers 제목 키워드 기반

Upmixing via style transfer: a variational autoencoder for disentangling spatial images and musical content

2022-03-22 · Haici Yang, Sanna Wager, Spencer Russell, Mike Luo 외

In the stereo-to-multichannel upmixing problem for music, one of the main tasks is to set the directionality of the instrument sources in the multichannel rendering results. In this paper, we propose a modified variation…

Style Transfer

Music Style Transfer With Diffusion Model

2024-04-23 · Hong Huang, Yuyi Wang, Luyao Li, Jun Lin

Previous studies on music style transfer have mainly focused on one-to-one style conversion, which is relatively limited. When considering the conversion between multiple styles, previous methods required designing multi…

Audio GenerationmodelMusic Style TransferStyle Transfer

Self-Supervised VQ-VAE for One-Shot Music Style Transfer

2021-02-10 · Ondřej Cífka, Alexey Ozerov, Umut Şimşekli, Gaël Richard

Neural style transfer, allowing to apply the artistic style of one image to another, has become one of the most widely showcased computer vision applications shortly after its introduction. In contrast, related tasks in …

Music Style TransferSelf-Supervised LearningStyle Transfer

Learning Interpretable Representation for Controllable Polyphonic Music Generation

2020-08-17 · Ziyu Wang, Dingsu Wang, Yixiao Zhang, Gus Xia

While deep generative models have become the leading methods for algorithmic composition, it remains a challenging problem to control the generation process because the latent variables of most deep-learning models lack …

DisentanglementMusic GenerationStyle Transfer

Boosting Multi-Speaker Expressive Speech Synthesis with Semi-supervised Contrastive Learning

2023-10-26 · Xinfa Zhu, Yuke Li, Yi Lei, Ning Jiang 외

This paper aims to build a multi-speaker expressive TTS system, synthesizing a target speaker's speech with multiple styles and emotions. To this end, we propose a novel contrastive learning-based TTS approach to transfe…

Contrastive LearningExpressive Speech SynthesisSpeech Synthesis