paper-with-me

홈 › Papers

Raw Multi-Channel Audio Source Separation using Multi-Resolution Convolutional Auto-Encoders

2018-03-02 · Emad M. Grais, Dominic Ward, Mark D. Plumbley

Supervised multi-channel audio source separation requires extracting useful spectral, temporal, and spatial features from the mixed signals. The success of many existing systems is therefore largely dependent on the choice of features used for training. In this work, we introduce a novel multi-channel, multi-resolution convolutional auto-encoder neural network that works on raw time-domain signals to determine appropriate multi-resolution features for separating the singing-voice from stereo music. Our experimental results show that the proposed method can achieve multi-channel audio source separation without the need for hand-crafted features or any pre- or post-processing.

📄 PDF Abstract BibTeX arXiv:1803.00702

Code (0)

등록된 구현이 없습니다.

Tasks

Audio Source Separation

Similar Papers 제목 키워드 기반

Independent Deeply Learned Matrix Analysis for Multichannel Audio Source Separation

2018-06-27

In this paper, we address a multichannel audio source separation task and propose a new efficient method called independent deeply learned matrix analysis (IDLMA). IDLMA estimates the demixing matrix in a blind manner an…

Audio Source Separation

A comprehensive study of speech separation: spectrogram vs waveform separation

2019-05-17 · Fahimeh Bahmaninezhad, Jian Wu, Rongzhi Gu, Shi-Xiong Zhang 외

Speech separation has been studied widely for single-channel close-talk microphone recordings over the past few years; developed solutions are mostly in frequency-domain. Recently, a raw audio waveform separation network…

speech-recognitionSpeech RecognitionSpeech Separation

DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification

2024-09-19 · Dongheon Lee, Jung-Woo Choi

This paper presents a framework for universal sound separation and polyphonic audio classification, addressing the challenges of separating and classifying individual sound sources in a multichannel mixture. The proposed…

Audio ClassificationClassificationMamba

On Neural Architectures for Deep Learning-based Source Separation of Co-Channel OFDM Signals

2023-03-11 · Gary C. F. Lee, Amir Weiss, Alejandro Lancho, Yury Polyanskiy 외

We study the single-channel source separation problem involving orthogonal frequency-division multiplexing (OFDM) signals, which are ubiquitous in many modern-day digital communication systems. Related efforts have been …

Time SeriesTime Series Analysis

Audio-visual Multi-channel Integration and Recognition of Overlapped Speech

2020-11-16 · Jianwei Yu, Shi-Xiong Zhang, Bo Wu, Shansong Liu 외

Automatic speech recognition (ASR) technologies have been significantly advanced in the past few decades. However, recognition of overlapped speech remains a highly challenging task to date. To this end, multi-channel mi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1