paper-with-me

홈 › Papers

Music FaderNets: Controllable Music Generation Based On High-Level Features via Low-Level Feature Modelling

2020-07-29 · Hao Hao Tan, Dorien Herremans

High-level musical qualities (such as emotion) are often abstract, subjective, and hard to quantify. Given these difficulties, it is not easy to learn good feature representations with supervised learning techniques, either because of the insufficiency of labels, or the subjectiveness (and hence large variance) in human-annotated labels. In this paper, we present a framework that can learn high-level feature representations with a limited amount of data, by first modelling their corresponding quantifiable low-level attributes. We refer to our proposed framework as Music FaderNets, which is inspired by the fact that low-level attributes can be continuously manipulated by separate "sliding faders" through feature disentanglement and latent regularization techniques. High-level features are then inferred from the low-level representations through semi-supervised clustering using Gaussian Mixture Variational Autoencoders (GM-VAEs). Using arousal as an example of a high-level feature, we show that the "faders" of our model are disentangled and change linearly w.r.t. the modelled low-level attributes of the generated output music. Furthermore, we demonstrate that the model successfully learns the intrinsic relationship between arousal and its corresponding low-level attributes (rhythm and note density), with only 1% of the training set being labelled. Finally, using the learnt high-level feature representations, we explore the application of our framework in style transfer tasks across different arousal states. The effectiveness of this approach is verified through a subjective listening test.

📄 PDF Abstract BibTeX arXiv:2007.15474

Code (1)

gudgud96/music-fader-nets 공식 구현 pytorch

Tasks

ClusteringDisentanglementMusic GenerationRhythmStyle Transfer

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

XMusic: Towards a Generalized and Controllable Symbolic Music Generation Framework

2025-01-15 · Sida Tian, Can Zhang, Wei Yuan, Wei Tan 외

In recent years, remarkable advancements in artificial intelligence-generated content (AIGC) have been achieved in the fields of image synthesis and text generation, generating content comparable to that produced by huma…

Emotion RecognitionImage GenerationMulti-Task LearningMusic Generation+1

CoCoFormer: A controllable feature-rich polyphonic music generation method

2023-10-15 · Jiuyang Zhou, Tengfei Niu, Hong Zhu, Xingping Wang

This paper explores the modeling method of polyphonic music sequence. Due to the great potential of Transformer models in music generation, controllable music generation is receiving more attention. In the task of polyph…

DiversityMusic GenerationRhythm

Controllable Lyrics-to-Melody Generation

2023-06-05 · Zhe Zhang, Yi Yu, Atsuhiro Takasu

Lyrics-to-melody generation is an interesting and challenging topic in AI music research field. Due to the difficulty of learning the correlations between lyrics and melody, previous methods suffer from low generation qu…

Music Generation

Vis2Mus: Exploring Multimodal Representation Mapping for Controllable Music Generation

2022-11-10 · Runbang Zhang, Yixiao Zhang, Kai Shao, Ying Shan 외

In this study, we explore the representation mapping from the domain of visual arts to the domain of music, with which we can use visual arts as an effective handle to control music generation. Unlike most studies in mul…

Music GenerationRepresentation LearningStyle Transfer

FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control

2022-01-26 · Dimitri von Rütte, Luca Biggio, Yannic Kilcher, Thomas Hofmann

Generating music with deep neural networks has been an area of active research in recent years. While the quality of generated samples has been steadily increasing, most methods are only able to exert minimal control ove…

Inductive BiasMusic Generation