paper-with-me

홈 › Papers

Modeling Melodic Feature Dependency with Modularized Variational Auto-Encoder

2018-10-31 · Yu-An Wang, Yu-Kai Huang, Tzu-Chuan Lin, Shang-Yu Su, Yun-Nung Chen

Automatic melody generation has been a long-time aspiration for both AI researchers and musicians. However, learning to generate euphonious melodies has turned out to be highly challenging. This paper introduces 1) a new variant of variational autoencoder (VAE), where the model structure is designed in a modularized manner in order to model polyphonic and dynamic music with domain knowledge, and 2) a hierarchical encoding/decoding strategy, which explicitly models the dependency between melodic features. The proposed framework is capable of generating distinct melodies that sounds natural, and the experiments for evaluating generated music clips show that the proposed model outperforms the baselines in human evaluation.

📄 PDF Abstract BibTeX arXiv:1811.00162

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

STS Classification with Dual-stream CNN

2018-05-20 · Shuchen Weng, Wenbo Li, Yi Zhang, Siwei Lyu

The structured time series (STS) classification problem requires the modeling of interweaved spatiotemporal dependency. most previous STS classification methods model the spatial and temporal dependencies independently. …

Activity RecognitionClassificationGeneral ClassificationSTS+2

Melodic Contour and Mid-Level Global Features Applied to the Analysis of Flamenco Cantes

2015-09-16 · Gómez Francisco, Mora Joaquín, Gómez Emilia, Díaz-Báñez José Miguel

This work focuses on the topic of melodic characterization and similarity in a specific musical repertoire: a cappella flamenco singing, more specifically in debla and martinete styles. We propose the combination of manu…

Hierarchical Generative Modeling of Melodic Vocal Contours in Hindustani Classical Music

2024-08-22 · Nithya Shikarpur, Krishna Maneesha Dendukuri, Yusong Wu, Antoine Caillon 외

Hindustani music is a performance-driven oral tradition that exhibits the rendition of rich melodic patterns. In this paper, we focus on generative modeling of singers' vocal melodies extracted from audio recordings, as …

Audio Synthesis

Semi-Implicit Graph Variational Auto-Encoders

2019-08-19 · NeurIPS 2019 12 · Arman Hasanzadeh, Ehsan Hajiramezanali, Nick Duffield, Krishna R. Narayanan 외

Semi-implicit graph variational auto-encoder (SIG-VAE) is proposed to expand the flexibility of variational graph auto-encoders (VGAE) to model graph data. SIG-VAE employs a hierarchical variational framework to enable n…

DecoderVariational Inference

Variational Autoencoders with Jointly Optimized Latent Dependency Structure

2019-05-01 · ICLR 2019 5 · Jiawei He, Yu Gong, Joseph Marino, Greg Mori 외

We propose a method for learning the dependency structure between latent variables in deep latent variable models. Our general modeling and inference framework combines the complementary strengths of deep generative mod…