Multiple Speaker Separation from Noisy Sources in Reverberant Rooms using Relative Transfer Matrix
Separation of simultaneously active multiple speakers is a difficult task in environments with strong reverberation and many background noise sources. This paper uses the relative transfer matrix (ReTM), a generalization of the relative transfer function of a room, to propose a simple yet novel approach for separating concurrent speakers using noisy multichannel microphone recordings. The proposed method (i) allows multiple speech and background noise sources, (ii) includes reverberation, (iii) does not need the knowledge of the locations of speech and noise sources nor microphone locations and their relative geometry, and (iv) uses relatively small segment of recordings for training. We illustrate the speech source separation capability with improved intelligibility using a simulation study consisting of four speakers in the presence of three noise sources in a reverberant room. We also show the applicability of the method in a practical experiment in a real room.
Code (0)
등록된 구현이 없습니다.
Tasks
Speaker SeparationSimilar Papers 제목 키워드 기반
Single channel voice separation for unknown number of speakers under reverberant and noisy settings
We present a unified network for voice separation of an unknown number of speakers. The proposed approach is composed of several separation heads optimized together with a speaker classification branch. The separation is…
ClassificationGeneral ClassificationAttractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
This paper addresses the problem of single-channel speech separation, where the number of speakers is unknown, and each speaker may speak multiple utterances. We propose a speech separation model that simultaneously perf…
Speech SeparationTime-Domain Speech Extraction with Spatial Information and Multi Speaker Conditioning Mechanism
In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on a…
Speech Extractionspeech-recognitionSpeech RecognitionSpeech SeparationTowards Real-Time Single-Channel Speech Separation in Noisy and Reverberant Environments
Real-time single-channel speech separation aims to unmix an audio stream captured from a single microphone that contains multiple people talking at once, environmental noise, and reverberation into multiple de-reverberat…
DecoderSpeech SeparationSingle-Microphone Speaker Separation and Voice Activity Detection in Noisy and Reverberant Environments
Speech separation involves extracting an individual speaker's voice from a multi-speaker audio signal. The increasing complexity of real-world environments, where multiple speakers might converse simultaneously, undersco…
Action DetectionActivity DetectionDecoderSpeaker Separation+1