paper-with-me

Papers

Deep Transform: Cocktail Party Source Separation via Probabilistic Re-Synthesis

2015-03-20 · Andrew J. R. Simpson

In cocktail party listening scenarios, the human brain is able to separate competing speech signals. However, the signal processing implemented by the brain to perform cocktail party listening is not well understood. Here, we trained two separate convolutive autoencoder deep neural networks (DNN) to separate monaural and binaural mixtures of two concurrent speech streams. We then used these DNNs as convolutive deep transform (CDT) devices to perform probabilistic re-synthesis. The CDTs operated directly in the time-domain. Our simulations demonstrate that very simple neural networks are capable of exploiting monaural and binaural information available in a cocktail party listening scenario.

📄 PDF Abstract BibTeX arXiv:1503.06046

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Deep Transform: Cocktail Party Source Separation via Complex Convolution in a Deep Neural Network

2015-04-12 · Andrew J. R. Simpson

Convolutional deep neural networks (DNN) are state of the art in many engineering problems but have not yet addressed the issue of how to deal with complex spectrograms. Here, we use circular statistics to provide a conv…

Probabilistic Binary-Mask Cocktail-Party Source Separation in a Convolutional Deep Neural Network

2015-03-24 · Andrew J. R. Simpson

Separation of competing speech is a key challenge in signal processing and a feat routinely performed by the human auditory brain. A long standing benchmark of the spectrogram approach to source separation is known as th…

Prediction

Permutation Invariant Training of Deep Models for Speaker-Independent Multi-talker Speech Separation

2016-07-01 · Dong Yu, Morten Kolbæk, Zheng-Hua Tan, Jesper Jensen

We propose a novel deep learning model, which supports permutation invariant training (PIT), for speaker independent multi-talker speech separation, commonly known as the cocktail-party problem. Different from most of th…

ClusteringDeep ClusteringDeep Learningregression+1

Deep Karaoke: Extracting Vocals from Musical Mixtures Using a Convolutional Deep Neural Network

2015-04-17 · Andrew J. R. Simpson, Gerard Roma, Mark D. Plumbley

Identification and extraction of singing voice from within musical mixtures is a key challenge in source separation and machine audition. Recently, deep neural networks (DNN) have been used to estimate 'ideal' binary mas…

Speech Separation

Cocktail Party Processing via Structured Prediction

2012-12-01 · NeurIPS 2012 12 · Yuxuan Wang, DeLiang Wang

While human listeners excel at selectively attending to a conversation in a cocktail party, machine performance is still far inferior by comparison. We show that the cocktail party problem, or the speech separation probl…

General ClassificationPredictionSpeech SeparationStructured Prediction