paper-with-me

Papers Music Modeling

“Music Modeling” 태그가 달린 논문 38편 · 필터 해제

Exploring LLMs for South Asian Music Understanding and Generation

2026-06-03 · Faria Binte Kader, Mohtasim Hadi Rafi, Shah Wasif Sajjad, Santu Karmaker arxiv

Recent advancements in Large Language Models (LLMs) have shown promising results in music understanding and generation tasks. However, existing works remain confined to Western tonal traditions, offering little insight i…

Music GenerationMusic Modeling

Depth-Structured Music Recurrence: Budgeted Recurrent Attention for Full-Piece Symbolic Music Modeling

2026-02-23 · Yungang Yi, Weihua Li, Matthew Kuo, Catherine Shi 외 arxiv

Long-context modeling is essential for symbolic music generation, since motif repetition and developmental variation can span thousands of musical events, yet practical workflows frequently rely on resource-limited hardw…

Music GenerationMusic Modeling

Pianist Transformer: Towards Expressive Piano Performance Rendering via Scalable Self-Supervised Pre-Training

2025-12-02 · Hong-Jie You, Jie-Jing Shao, Xiao-Wen Yang, Lin-Han Jia 외 arxiv

Existing methods for expressive music performance rendering, a conditional generation task that aims to generate a human-like performance from a symbolic score, rely on supervised learning over small labeled datasets, wh…

Self-Supervised LearningMusic Modeling

HNote: Extending YNote with Hexadecimal Encoding for Fine-Tuning LLMs in Music Modeling

2025-09-30 · Hung-Ying Chu, Shao-Yu Wei, Guan-Wei Chen, Tzu-Wei Hung 외 arxiv

Recent advances in large language models (LLMs) have created new opportunities for symbolic music generation. However, existing formats such as MIDI, ABC, and MusicXML are either overly complex or structurally inconsiste…

Music GenerationMusic Modeling

SLEEPING-DISCO 9M: A large-scale pre-training dataset for generative music modeling

2025-06-17 · Tawsif Ahmed, Andrej Radonjic, Gollam Rabby

We present Sleeping-DISCO 9M, a large-scale pre-training dataset for music and song. To the best of our knowledge, there are no open-source high-quality dataset representing popular and well-known songs for generative mu…

Music CaptioningMusic ModelingSinging Voice Synthesis

Frechet Music Distance: A Metric For Generative Symbolic Music Evaluation

2024-12-10 · Jan Retkowski, Jakub Stępniak, Mateusz Modrzejewski

In this paper we introduce the Frechet Music Distance (FMD), a novel evaluation metric for generative symbolic music models, inspired by the Frechet Inception Distance (FID) in computer vision and Frechet Audio Distance …

FADMusic GenerationMusic Modeling

MuPT: A Generative Symbolic Music Pretrained Transformer

2024-04-09 · Xingwei Qu, Yuelin Bai, Yinghao Ma, Ziya Zhou 외

In this paper, we explore the application of Large Language Models (LLMs) to the pre-training of music. While the prevalent use of MIDI in music modeling is well-established, our findings suggest that LLMs are inherently…

Music GenerationMusic Modeling

Impact of time and note duration tokenizations on deep learning symbolic music modeling

2023-10-12 · Nathan Fradet, Nicolas Gutowski, Fabien Chhel, Jean-Pierre Briot

Symbolic music is widely used in various deep learning tasks, including generation, transcription, synthesis, and Music Information Retrieval (MIR). It is mostly employed with discrete models like Transformers, which req…

Emotion ClassificationInformation RetrievalMusic GenerationMusic Information Retrieval+3

A Domain-Knowledge-Inspired Music Embedding Space and a Novel Attention Mechanism for Symbolic Music Modeling

2022-12-02 · Z. Guo, J. Kang, D. Herremans

Following the success of the transformer architecture in the natural language domain, transformer-like architectures have been widely applied to the domain of symbolic music recently. Symbolic music and text, however, ar…

Music GenerationMusic Modeling

Low-Rank Constraints for Fast Inference in Structured Models

2022-01-08 · NeurIPS 2021 12 · Justin T. Chiu, Yuntian Deng, Alexander M. Rush

Structured distributions, i.e. distributions over combinatorial spaces, are commonly used to learn latent probabilistic representations from observed data. However, scaling these models is bottlenecked by the high comput…

Language ModelingLanguage ModellingMusic Modeling

Gates Are Not What You Need in RNNs

2021-08-01 · Ronalds Zakovskis, Andis Draguns, Eliza Gaile, Emils Ozolins 외

Recurrent neural networks have flourished in many areas. Consequently, we can see new RNN cells being developed continuously, usually by creating or using gates in a new, original way. But what if we told you that gates …

Language ModelingLanguage ModellingMusic ModelingSentiment Analysis

MuSLCAT: Multi-Scale Multi-Level Convolutional Attention Transformer for Discriminative Music Modeling on Raw Waveforms

2021-04-06 · Kai Middlebrook, Shyam Sudhakaran, David Guy Brizan

In this work, we aim to improve the expressive capacity of waveform-based discriminative music networks by modeling both sequential (temporal) and hierarchical information in an efficient end-to-end architecture. We pres…

Music ModelingMusic Tagging

Rethinking Neural Operations for Diverse Tasks

2021-03-29 · NeurIPS 2021 12 · Nicholas Roberts, Mikhail Khodak, Tri Dao, Liam Li 외

An important goal of AutoML is to automate-away the design of neural networks on new tasks in under-explored domains. Motivated by this goal, we study the problem of enabling users to discover the right neural operations…

AutoMLImage ClassificationInductive BiasMusic Modeling+2

Recurrently Controlling a Recurrent Network with Recurrent Networks Controlled by More Recurrent Networks

2021-01-01 · Yi Tay, Yikang Shen, Alvin Chan, Aston Zhang 외

This paper explores an intriguing idea of recursively parameterizing recurrent nets. Simply speaking, this refers to recurrently controlling a recurrent network with recurrent networks controlled by recurrent networks. T…

Code GenerationInductive BiasMachine TranslationMusic Modeling+2

PopMAG: Pop Music Accompaniment Generation

2020-08-18 · Yi Ren, Jinzheng He, Xu Tan, Tao Qin 외

In pop music, accompaniments are usually played by multiple instruments (tracks) such as drum, bass, string and guitar, and can make a song more expressive and contagious by arranging together with its melody. Previous w…

Music Modeling

Pop Music Transformer: Beat-based Modeling and Generation of Expressive Pop Piano Compositions

2020-02-01 · Yu-Siang Huang, Yi-Hsuan Yang

A great number of deep learning based models have been recently proposed for automatic music composition. Among these models, the Transformer stands out as a prominent approach for generating expressive classical piano p…

Music Modeling

Learning Style-Aware Symbolic Music Representations by Adversarial Autoencoders

2020-01-15 · Andrea Valenti, Antonio Carta, Davide Bacciu

We address the challenging open problem of learning an effective latent space for symbolic music data in generative music modeling. We focus on leveraging adversarial regularization as a flexible and natural mean to imbu…

Music Modeling

Metagross: Meta Gated Recursive Controller Units for Sequence Modeling

2020-01-01 · ICLR 2020 1 · Yi Tay, Yikang Shen, Alvin Chan, Yew Soon Ong

This paper proposes Metagross (Meta Gated Recursive Controller), a new neural sequence modeling unit. Our proposed unit is characterized by recursive parameterization of its gating functions, i.e., gating mechanisms of M…

Code GenerationInductive BiasMachine TranslationMusic Modeling+2

Improving Polyphonic Music Models with Feature-Rich Encoding

2019-11-26 · Omar Peracha

This paper explores sequential modelling of polyphonic music with deep neural networks. While recent breakthroughs have focussed on network architecture, we demonstrate that the representation of the sequence can make an…

Music GenerationMusic Modeling

Gating Revisited: Deep Multi-layer RNNs That Can Be Trained

2019-11-25 · Mehmet Ozgur Turkoglu, Stefano D'Aronco, Jan Dirk Wegner, Konrad Schindler

We propose a new STAckable Recurrent cell (STAR) for recurrent neural networks (RNNs), which has fewer parameters than widely used LSTM and GRU while being more robust against vanishing or exploding gradients. Stacking r…

Action RecognitionAction Recognition In VideosLanguage ModellingMusic Modeling+1
1–20 / 38 다음 →