paper-with-me

홈 › Papers

Story2MIDI: Emotionally Aligned Music Generation from Text

2025-12-01 · Mohammad Shokri, Alexandra C. Salem, Gabriel Levine, Johanna Devaney, Sarah Ita Levitan arxiv

In this paper, we introduce Story2MIDI, a sequence-to-sequence Transformer-based model for generating emotion-aligned music from a given piece of text. To develop this model, we construct the Story2MIDI dataset by merging existing datasets for sentiment analysis from text and emotion classification in music. The resulting dataset contains pairs of text blurbs and music pieces that evoke the same emotions in the reader or listener. Despite the small scale of our dataset and limited computational resources, our results indicate that our model effectively learns emotion-relevant features in music and incorporates them into its generation process, producing samples with diverse emotional responses. We evaluate the generated outputs using objective musical metrics and a human listening study, confirming the model's ability to capture intended emotional cues.

📄 PDF Abstract BibTeX arXiv:2512.02192

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion ClassificationSentiment AnalysisMusic Generation

Similar Papers 제목 키워드 기반

Emotion-Guided Image to Music Generation

2024-10-29 · Souraja Kundu, Saket Singh, Yuji Iwahori

Generating music from images can enhance various applications, including background music for photo slideshows, social media experiences, and video creation. This paper presents an emotion-guided image-to-music generatio…

Contrastive LearningMusic Generation

Video-based Music Generation

2026-02-05 · Serkan Sulun arxiv

As the volume of video content on the internet grows rapidly, finding a suitable soundtrack remains a significant challenge. This thesis presents EMSYNC (EMotion and SYNChronization), a fast, free, and automatic solution…

Emotion ClassificationMusic Generation

XMusic: Towards a Generalized and Controllable Symbolic Music Generation Framework

2025-01-15 · Sida Tian, Can Zhang, Wei Yuan, Wei Tan 외

In recent years, remarkable advancements in artificial intelligence-generated content (AIGC) have been achieved in the fields of image synthesis and text generation, generating content comparable to that produced by huma…

Emotion RecognitionImage GenerationMulti-Task LearningMusic Generation+1

BERT-like Pre-training for Symbolic Piano Music Classification Tasks

2021-07-12 · Yi-Hui Chou, I-Chun Chen, Chin-Jui Chang, Joann Ching 외

This article presents a benchmark study of symbolic piano music classification using the masked language modelling approach of the Bidirectional Encoder Representations from Transformers (BERT). Specifically, we consider…

ClassificationEmotion ClassificationLanguage ModellingMelody Extraction+1

Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment

2025-05-19 · Abhinaba Roy, Geeta Puri, Dorien Herremans

We present Text2midi-InferAlign, a novel technique for improving symbolic music generation at inference time. Our method leverages text-to-audio alignment and music structural alignment rewards during inference to encour…

Music Generation