paper-with-me

홈 › Papers

Interpreting Graphic Notation with MusicLDM: An AI Improvisation of Cornelius Cardew's Treatise

2024-12-12 · Tornike Karchkhadze, Keren Shao, Shlomo Dubnov

This work presents a novel method for composing and improvising music inspired by Cornelius Cardew's Treatise, using AI to bridge graphic notation and musical expression. By leveraging OpenAI's ChatGPT to interpret the abstract visual elements of Treatise, we convert these graphical images into descriptive textual prompts. These prompts are then input into MusicLDM, a pre-trained latent diffusion model designed for music generation. We introduce a technique called "outpainting," which overlaps sections of AI-generated music to create a seamless and cohesive composition. We demostrate a new perspective on performing and interpreting graphic scores, showing how AI can transform visual stimuli into sound and expand the creative possibilities in contemporary/experimental music composition. Musical pieces are available at https://bit.ly/TreatiseAI

📄 PDF Abstract BibTeX arXiv:2412.08944

Code (0)

등록된 구현이 없습니다.

Tasks

DescriptiveMusic Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.

Similar Papers 제목 키워드 기반

MusicLDM: Enhancing Novelty in Text-to-Music Generation Using Beat-Synchronous Mixup Strategies

2023-08-03 · Ke Chen, Yusong Wu, Haohe Liu, Marianna Nezhurina 외

Diffusion models have shown promising results in cross-modal generation tasks, including text-to-image and text-to-audio generation. However, generating music, as a special type of audio, presents unique challenges due t…

Audio GenerationBeat TrackingData AugmentationMusic Generation+1

Music Interpretation and Emotion Perception: A Computational and Neurophysiological Investigation

2025-05-16 · Vassilis Lyberatos, Spyridon Kantarelis, Ioanna Zioga, Christina Anagnostopoulou 외

This study investigates emotional expression and perception in music performance using computational and neurophysiological methods. The influence of different performance settings, such as repertoire, diatonic modal etu…

Emotion Recognition

Emovectors: assessing emotional content in jazz improvisations for creativity evaluation

2025-12-09 · Anna Jordanous arxiv

Music improvisation is fascinating to study, being essentially a live demonstration of a creative process. In jazz, musicians often improvise across predefined chord progressions (leadsheets). How do we assess the creati…

Detecting and Characterizing Planning in Language Models

2025-08-25 · Jatin Nainani, Sankaran Vaidyanathan, Connor Watts, Andre N. Assis 외 arxiv

Modern large language models (LLMs) have demonstrated impressive performance across a wide range of multi-step reasoning tasks. Recent work suggests that LLMs may perform planning - selecting a future target token in adv…

Code Generation

CALLIG: Computer Assisted Language Learning using Improvisation Games

2020-05-01 · LREC 2020 5 · Lu{\'\i}s Morgado da Costa, Joanna Ut-Seong Sio

In this paper, we present the ongoing development of CALLIG {--} a web system that uses improvisation games in Computer Assisted Language Learning (CALL). Improvisation games are structured activities with built-in const…