Audio-based Distributional Semantic Models for Music Auto-tagging and Similarity Measurement
The recent development of Audio-based Distributional Semantic Models (ADSMs) enables the computation of audio and lexical vector representations in a joint acoustic-semantic space. In this work, these joint representations are applied to the problem of automatic tag generation. The predicted tags together with their corresponding acoustic representation are exploited for the construction of acoustic-semantic clip embeddings. The proposed algorithms are evaluated on the task of similarity measurement between music clips. Acoustic-semantic models are shown to outperform the state-of-the-art for this task and produce high quality tags for audio/music clips.
Code (0)
등록된 구현이 없습니다.
Tasks
Music Auto-TaggingTAGSimilar Papers 제목 키워드 기반
Joint Music and Language Attention Models for Zero-shot Music Tagging
Music tagging is a task to predict the tags of music recordings. However, previous music tagging research primarily focuses on close-set music tagging tasks which can not be generalized to new tags. In this work, we prop…
Audio TaggingDecoderMusic TaggingZero-shot Learning for Audio-based Music Classification and Tagging
Audio-based music classification and tagging is typically based on categorical supervised learning with a fixed set of labels. This intrinsically cannot handle unseen labels such as newly added music genres or semantic w…
AttributeClassificationGeneral ClassificationMulti-label zero-shot learning+4A Deep Bag-of-Features Model for Music Auto-Tagging
Feature learning and deep learning have drawn great attention in recent years as a way of transforming input data into more effective representations using learning algorithms. Such interest has grown in the area of musi…
Audio ClassificationInformation RetrievalMusic Auto-TaggingMusic Information Retrieval+2Zero-shot Learning and Knowledge Transfer in Music Classification and Tagging
Music classification and tagging is conducted through categorical supervised learning with a fixed set of labels. In principle, this cannot make predictions on unseen labels. Zero-shot learning is an approach to solve th…
ClassificationGeneral ClassificationMusic ClassificationTransfer Learning+1Multi-Level and Multi-Scale Feature Aggregation Using Pre-trained Convolutional Neural Networks for Music Auto-tagging
Music auto-tagging is often handled in a similar manner to image classification by regarding the 2D audio spectrogram as image data. However, music auto-tagging is distinguished from image classification in that the tags…
General Classificationimage-classificationImage ClassificationMusic Auto-Tagging+1