paper-with-me

Papers

Raw Waveform-based Audio Classification Using Sample-level CNN Architectures

2017-12-04 · Jongpil Lee, Taejun Kim, Jiyoung Park, Juhan Nam

Music, speech, and acoustic scene sound are often handled separately in the audio domain because of their different signal characteristics. However, as the image domain grows rapidly by versatile image classification models, it is necessary to study extensible classification models in the audio domain as well. In this study, we approach this problem using two types of sample-level deep convolutional neural networks that take raw waveforms as input and uses filters with small granularity. One is a basic model that consists of convolution and pooling layers. The other is an improved model that additionally has residual connections, squeeze-and-excitation modules and multi-level concatenation. We show that the sample-level models reach state-of-the-art performance levels for the three different categories of sound. Also, we visualize the filters along layers and compare the characteristics of learned filters.

📄 PDF Abstract BibTeX arXiv:1712.00866

Code (0)

등록된 구현이 없습니다.

Tasks

Audio ClassificationClassificationGeneral Classificationimage-classificationImage Classification

Similar Papers 제목 키워드 기반

A Generative Model for Raw Audio Using Transformer Architectures

2021-06-30 · Prateek Verma, Chris Chafe

This paper proposes a novel way of doing audio synthesis at the waveform level using Transformer architectures. We propose a deep neural network for generating waveforms, similar to wavenet. This is fully probabilistic, …

Audio Synthesis

Sample-level CNN Architectures for Music Auto-tagging Using Raw Waveforms

2017-10-28 · Taejun Kim, Jongpil Lee, Juhan Nam

Recent work has shown that the end-to-end approach using convolutional neural network (CNN) is effective in various types of machine learning tasks. For audio signals, the approach takes raw waveforms as input using an 1…

General Classificationimage-classificationMusic Auto-Tagging

Audio Classification of Bit-Representation Waveform

2019-04-08 · Masaki Okawa, Takuya Saito, Naoki Sawada, Hiromitsu Nishizaki

This study investigated the waveform representation for audio signal classification. Recently, many studies on audio waveform classification such as acoustic event detection and music genre classification have been publi…

Audio ClassificationClassificationEvent DetectionGeneral Classification+3

An End-to-End Audio Classification System based on Raw Waveforms and Mix-Training Strategy

2019-11-21 · Jiaxu Chen, Jing Hao, Kai Chen, Di Xie 외

Audio classification can distinguish different kinds of sounds, which is helpful for intelligent applications in daily life. However, it remains a challenging task since the sound events in an audio clip is probably mult…

Audio ClassificationClassificationGeneral ClassificationMulti-Label Classification+1

Multi-Level and Multi-Scale Feature Aggregation Using Sample-level Deep Convolutional Neural Networks for Music Classification

2017-06-21 · Jongpil Lee, Juhan Nam

Music tag words that describe music audio by text have different levels of abstraction. Taking this issue into account, we propose a music classification approach that aggregates multi-level and multi-scale features usin…

ClassificationGeneral ClassificationMusic ClassificationTAG