paper-with-me

Papers

Integrating Text-to-Music Models with Language Models: Composing Long Structured Music Pieces

2024-10-01 · Lilac Atassi

Recent music generation methods based on transformers have a context window of up to a minute. The music generated by these methods is largely unstructured beyond the context window. With a longer context window, learning long-scale structures from musical data is a prohibitively challenging problem. This paper proposes integrating a text-to-music model with a large language model to generate music with form. The papers discusses the solutions to the challenges of such integration. The experimental results show that the proposed method can generate 2.5-minute-long music that is highly structured, strongly organized, and cohesive.

📄 PDF Abstract BibTeX arXiv:2410.00344

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelMusic Generation

Similar Papers 제목 키워드 기반

ComposeOn Academy: Transforming Melodic Ideas into Complete Compositions Integrating Music Learning

2025-02-21 · Hongxi Pu, Futian Jiang, Zihao Chen, Xingyue Song

Music composition has long been recognized as a significant art form. However, existing digital audio workstations and music production software often present high entry barriers for users lacking formal musical training…

JEN-1 Composer: A Unified Framework for High-Fidelity Multi-Track Music Generation

2023-10-29 · Yao Yao, Peike Li, BoYu Chen, Alex Wang

With rapid advances in generative artificial intelligence, the text-to-music synthesis task has emerged as a promising direction for music generation. Nevertheless, achieving precise control over multi-track generation r…

Music Generation

MusiLingo: Bridging Music and Text with Pre-trained Language Models for Music Captioning and Query Response

2023-09-15 · Zihao Deng, Yinghao Ma, Yudong Liu, Rongchen Guo 외

Large Language Models (LLMs) have shown immense potential in multimodal applications, yet the convergence of textual and musical domains remains not well-explored. To address this gap, we present MusiLingo, a novel syste…

Caption GenerationLanguage ModellingMusic Captioning

ChatMusician: Understanding and Generating Music Intrinsically with LLM

2024-02-25 · Ruibin Yuan, Hanfeng Lin, Yi Wang, Zeyue Tian 외

While Large Language Models (LLMs) demonstrate impressive capabilities in text generation, we find that their ability has yet to be generalized to music, humanity's creative language. We introduce ChatMusician, an open-s…

MMLUText Generation

InverseMV: Composing Piano Scores with a Convolutional Video-Music Transformer

2021-12-31 · Chin-Tung Lin, Mu Yang

Many social media users prefer consuming content in the form of videos rather than text. However, in order for content creators to produce videos with a high click-through rate, much editing is needed to match the footag…

Music Generation