paper-with-me

Papers

Learning Transposition-Invariant Interval Features from Symbolic Music and Audio

2018-06-21 · Stefan Lattner, Maarten Grachten, Gerhard Widmer

Many music theoretical constructs (such as scale types, modes, cadences, and chord types) are defined in terms of pitch intervals---relative distances between pitches. Therefore, when computer models are employed in music tasks, it can be useful to operate on interval representations rather than on the raw musical surface. Moreover, interval representations are transposition-invariant, valuable for tasks like audio alignment, cover song detection and music structure analysis. We employ a gated autoencoder to learn fixed-length, invertible and transposition-invariant interval representations from polyphonic music in the symbolic domain and in audio. An unsupervised training method is proposed yielding an organization of intervals in the representation space which is musically plausible. Based on the representations, a transposition-invariant self-similarity matrix is constructed and used to determine repeated sections in symbolic music and in audio, yielding competitive results in the MIREX task "Discovery of Repeated Themes and Sections".

📄 PDF Abstract BibTeX arXiv:1806.08236

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Learning Musical Relations using Gated Autoencoders

2017-08-17 · Stefan Lattner, Maarten Grachten, Gerhard Widmer

Music is usually highly structured and it is still an open question how to design models which can successfully learn to recognize and represent musical structure. A fundamental problem is that structurally related patte…

Open-Ended Question AnsweringRhythm

Evaluating Interval-based Tokenization for Pitch Representation in Symbolic Music Analysis

2025-01-08 · Dinh-Viet-Toan Le, Louis Bigo, Mikaela Keller

Symbolic music analysis tasks are often performed by models originally developed for Natural Language Processing, such as Transformers. Such models require the input data to be represented as sequences, which is achieved…

Learning Complex Basis Functions for Invariant Representations of Audio

2019-07-13 · Stefan Lattner, Monika Dörfler, Andreas Arzt

Learning features from data has shown to be more successful than using hand-crafted features for many machine learning tasks. In music information retrieval (MIR), features learned from windowed spectrograms are highly v…

Information RetrievalMusic Information RetrievalRetrieval

Music102: An $D_{12}$-equivariant transformer for chord progression accompaniment

2024-10-23 · Weiliang Luo

We present Music102, an advanced model built upon the Music101 prototype, aimed at enhancing chord progression accompaniment through a D12-equivariant transformer. Inspired by group theory and symbolic music structures, …

Music Generation

Symphony Generation with Permutation Invariant Language Model

2022-05-10 · Jiafeng Liu, Yuanliang Dong, Zehua Cheng, Xinran Zhang 외

In this work, we propose a permutation invariant language model, SymphonyNet, as a solution for symbolic symphony music generation. We propose a novel Multi-track Multi-instrument Repeatable (MMR) representation for symp…

Audio GenerationDecoderLanguage ModelingLanguage Modelling+3