paper-with-me

홈 › Papers

Deep Learning Based Phase Reconstruction for Speaker Separation: A Trigonometric Perspective

2018-11-22 · Zhong-Qiu Wang, Ke Tan, DeLiang Wang

This study investigates phase reconstruction for deep learning based monaural talker-independent speaker separation in the short-time Fourier transform (STFT) domain. The key observation is that, for a mixture of two sources, with their magnitudes accurately estimated and under a geometric constraint, the absolute phase difference between each source and the mixture can be uniquely determined; in addition, the source phases at each time-frequency (T-F) unit can be narrowed down to only two candidates. To pick the right candidate, we propose three algorithms based on iterative phase reconstruction, group delay estimation, and phase-difference sign prediction. State-of-the-art results are obtained on the publicly available wsj0-2mix and 3mix corpus.

📄 PDF Abstract BibTeX arXiv:1811.09010

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Separation

Similar Papers 제목 키워드 기반

End-to-End Speech Separation with Unfolded Iterative Phase Reconstruction

2018-04-26 · Zhong-Qiu Wang, Jonathan Le Roux, DeLiang Wang, John R. Hershey

This paper proposes an end-to-end approach for single-channel speaker-independent multi-speaker speech separation, where time-frequency (T-F) masking, the short-time Fourier transform (STFT), and its inverse are represen…

Speech Separation

Speaker-independent Speech Separation with Deep Attractor Network

2017-07-12 · Yi Luo, Zhuo Chen, Nima Mesgarani

Despite the recent success of deep learning for many speech processing tasks, single-microphone, speaker-independent speech separation remains challenging for two main reasons. The first reason is the arbitrary order of …

Deep LearningSpeech Separation

Phasebook and Friends: Leveraging Discrete Representations for Source Separation

2018-10-02 · Jonathan Le Roux, Gordon Wichern, Shinji Watanabe, Andy Sarroff 외

Deep learning based speech enhancement and source separation systems have recently reached unprecedented levels of quality, to the point that performance is reaching a new ceiling. Most systems rely on estimating the mag…

Speaker SeparationSpeech Enhancement

Audio-visual Speech Separation with Adversarially Disentangled Visual Representation

2020-11-29 · Peng Zhang, Jiaming Xu, Jing Shi, Yunzhe Hao 외

Speech separation aims to separate individual voice from an audio mixture of multiple simultaneous talkers. Although audio-only approaches achieve satisfactory performance, they build on a strategy to handle the predefin…

Speech Separation

MIMO-DBnet: Multi-channel Input and Multiple Outputs DOA-aware Beamforming Network for Speech Separation

2022-12-07 · Yanjie Fu, Haoran Yin, Meng Ge, Longbiao Wang 외

Recently, many deep learning based beamformers have been proposed for multi-channel speech separation. Nevertheless, most of them rely on extra cues known in advance, such as speaker feature, face image or directional in…

Speech Separation