paper-with-me

홈 › Papers

A Comprehensive Survey on Deep Music Generation: Multi-level Representations, Algorithms, Evaluations, and Future Directions

2020-11-13 · Shulei Ji, Jing Luo, Xinyu Yang

The utilization of deep learning techniques in generating various contents (such as image, text, etc.) has become a trend. Especially music, the topic of this paper, has attracted widespread attention of countless researchers.The whole process of producing music can be divided into three stages, corresponding to the three levels of music generation: score generation produces scores, performance generation adds performance characteristics to the scores, and audio generation converts scores with performance characteristics into audio by assigning timbre or generates music in audio format directly. Previous surveys have explored the network models employed in the field of automatic music generation. However, the development history, the model evolution, as well as the pros and cons of same music generation task have not been clearly illustrated. This paper attempts to provide an overview of various composition tasks under different music generation levels, covering most of the currently popular music generation tasks using deep learning. In addition, we summarize the datasets suitable for diverse tasks, discuss the music representations, the evaluation methods as well as the challenges under different levels, and finally point out several future directions.

📄 PDF Abstract BibTeX arXiv:2011.06801

Code (0)

등록된 구현이 없습니다.

Tasks

Audio GenerationMusic Generation

Similar Papers 제목 키워드 기반

A Survey of AI Music Generation Tools and Models

2023-08-24 · Yueyue Zhu, Jared Baca, Banafsheh Rekabdar, Reza Rawassizadeh

In this work, we provide a comprehensive survey of AI music generation tools, including both research projects and commercialized applications. To conduct our analysis, we classified music generation approaches into thre…

Music GenerationSurvey

Vision-to-Music Generation: A Survey

2025-03-27 · Zhaokai Wang, Chenxi Bao, Le Zhuo, Jingrui Han 외

Vision-to-music Generation, including video-to-music and image-to-music tasks, is a significant branch of multimodal artificial intelligence demonstrating vast application prospects in fields such as film scoring, short …

multimodal generationMusic GenerationSurvey

A Comprehensive Survey on Generative AI for Video-to-Music Generation

2025-02-18 · Shulei Ji, Songruoyao Wu, ZiHao Wang, Shuyu Li 외

The burgeoning growth of video-to-music generation can be attributed to the ascendancy of multimodal generative models. However, there is a lack of literature that comprehensively combs through the work in this field. To…

Music Generation

A Survey on Music Generation from Single-Modal, Cross-Modal, and Multi-Modal Perspectives

2025-04-01 · Shuyu Li, Shulei Ji, ZiHao Wang, Songruoyao Wu 외

Multi-modal music generation, using multiple modalities like text, images, and video alongside musical scores and audio as guidance, is an emerging research area with broad applications. This paper reviews this field, ca…

Music Generation

Applications and Advances of Artificial Intelligence in Music Generation:A Review

2024-09-03 · Yanxu Chen, Linshu Huang, Tian Gou

In recent years, artificial intelligence (AI) has made significant progress in the field of music generation, driving innovation in music creation and applications. This paper provides a systematic review of the latest r…

Audio GenerationMusic Generation