paper-with-me

Papers

R&B -- Rhythm and Brain: Cross-subject Decoding of Music from Human Brain Activity

2024-06-21 · Matteo Ferrante, Matteo Ciferri, Nicola Toschi

Music is a universal phenomenon that profoundly influences human experiences across cultures. This study investigates whether music can be decoded from human brain activity measured with functional MRI (fMRI) during its perception. Leveraging recent advancements in extensive datasets and pre-trained computational models, we construct mappings between neural data and latent representations of musical stimuli. Our approach integrates functional and anatomical alignment techniques to facilitate cross-subject decoding, addressing the challenges posed by the low temporal resolution and signal-to-noise ratio (SNR) in fMRI data. Starting from the GTZan fMRI dataset, where five participants listened to 540 musical stimuli from 10 different genres while their brain activity was recorded, we used the CLAP (Contrastive Language-Audio Pretraining) model to extract latent representations of the musical stimuli and developed voxel-wise encoding models to identify brain regions responsive to these stimuli. By applying a threshold to the association between predicted and actual brain activity, we identified specific regions of interest (ROIs) which can be interpreted as key players in music processing. Our decoding pipeline, primarily retrieval-based, employs a linear map to project brain activity to the corresponding CLAP features. This enables us to predict and retrieve the musical stimuli most similar to those that originated the fMRI data. Our results demonstrate state-of-the-art identification accuracy, with our methods significantly outperforming existing approaches. Our findings suggest that neural-based music retrieval systems could enable personalized recommendations and therapeutic applications. Future work could use higher temporal resolution neuroimaging and generative models to improve decoding accuracy and explore the neural underpinnings of music perception and emotion.

📄 PDF Abstract BibTeX arXiv:2406.15537

Code (1)

neoayanami/fmri-music-retrieve 공식 구현 pytorch

Tasks

RetrievalRhythm

Similar Papers 제목 키워드 기반

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

2026-05-08 · Maryam Maghsoudi, Shihab Shamma arxiv

Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across subjects and sessions In this work, we propose a new approach to th…

Towards the bio-personalization of music recommendation systems: A single-sensor EEG biomarker of subjective music preference

2016-09-21 · Dimitrios A. Adamos, Stavros I. Dimitriadis, Nikolaos A. Laskaris

Recent advances in biosensors technology and mobile electroencephalographic (EEG) interfaces have opened new application fields for cognitive monitoring. A computable biomarker for the assessment of spontaneous aesthetic…

Brain Computer InterfaceEEGElectroencephalogram (EEG)Music Recommendation+1

A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses

2025-12-03 · Maryam Maghsoudi, Mohsen Rezaeizadeh, Shihab Shamma arxiv

Decoding imagined speech engages complex neural processes that are difficult to interpret due to uncertainty in timing and the limited availability of imagined-response datasets. In this study, we present a Magnetoenceph…

Modeling Temporal Lobe Epilepsy during Music Large-Scale Form Perception using the Impulse Pattern Formulation (IPF) Brain Mode

2023-10-05 · Rolf Bader

Musical large-scale form is investigated using an Electronic Dance Music (EDM) piece fed into a Finite-Difference Time Domain (FDTD) physical model of the cochlear which again inputs into an Impulse-Pattern Formulation (…

EEGFormRhythm

Diff-V2M: A Hierarchical Conditional Diffusion Model with Explicit Rhythmic Modeling for Video-to-Music Generation

2025-11-12 · Shulei Ji, Zihao Wang, Jiaxing Yu, Xiangyuan Yang 외 arxiv

Video-to-music (V2M) generation aims to create music that aligns with visual content. However, two main challenges persist in existing methods: (1) the lack of explicit rhythm modeling hinders audiovisual temporal alignm…

Music Generation