Learning Contextualized Music Semantics from Tags via a Siamese Network
Music information retrieval faces a challenge in modeling contextualized musical concepts formulated by a set of co-occurring tags. In this paper, we investigate the suitability of our recently proposed approach based on a Siamese neural network in fighting off this challenge. By means of tag features and probabilistic topic models, the network captures contextualized semantics from tags via unsupervised learning. This leads to a distributed semantics space and a potential solution to the out of vocabulary problem which has yet to be sufficiently addressed. We explore the nature of the resultant music-based semantics and address computational needs. We conduct experiments on three public music tag collections -namely, CAL500, MagTag5K and Million Song Dataset- and compare our approach to a number of state-of-the-art semantics learning approaches. Comparative results suggest that this approach outperforms previous approaches in terms of semantic priming and music tag completion.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalMusic Information RetrievalRetrievalTAGTopic ModelsSimilar Papers 제목 키워드 기반
Learning Contextualized Semantics from Co-occurring Terms via a Siamese Architecture
One of the biggest challenges in Multimedia information retrieval and understanding is to bridge the semantic gap by properly modeling concept semantics in context. The presence of out of vocabulary (OOV) concepts exacer…
DescriptiveInformation RetrievalRetrievalTopic ModelsTowards Deep Modeling of Music Semantics using EEG Regularizers
Modeling of music audio semantics has been previously tackled through learning of mappings from audio data to high-level tags or latent unsupervised spaces. The resulting semantic spaces are theoretically limited, either…
Cross-Modal RetrievalEEGElectroencephalogram (EEG)Retrieval+1Learning Contextual Tag Embeddings for Cross-Modal Alignment of Audio and Tags
Self-supervised audio representation learning offers an attractive alternative for obtaining generic audio embeddings, capable to be employed into various downstream tasks. Published approaches that consider both audio a…
cross-modal alignmentRepresentation LearningTAGWord EmbeddingsAutomatic Generation of Social Tags for Music Recommendation
Social tags are user-generated keywords associated with some resource on the Web. In the case of music, social tags have become an important component of Web2.0" recommender systems, allowing users to generate playlists …
Music RecommendationRecommendation SystemsTAGTag2Risk: Harnessing Social Music Tags for Characterizing Depression Risk
Musical preferences have been considered a mirror of the self. In this age of Big Data, online music streaming services allow us to capture ecologically valid music listening behavior and provide a rich source of informa…
valid