paper-with-me

홈 › Papers

Jointly Tracking and Separating Speech Sources Using Multiple Features and the generalized labeled multi-Bernoulli Framework

2018-04-16

This paper proposes a novel joint multi-speaker tracking-and-separation method based on the generalized labeled multi-Bernoulli (GLMB) multi-target tracking filter, using sound mixtures recorded by microphones. Standard multi-speaker tracking algorithms usually only track speaker locations, and ambiguity occurs when speakers are spatially close. The proposed multi-feature GLMB tracking filter treats the set of vectors of associated speaker features (location, pitch and sound) as the multi-target multi-feature observation, characterizes transitioning features with corresponding transition models and overall likelihood function, thus jointly tracks and separates each multi-feature speaker, and addresses the spatial ambiguity problem. Numerical evaluation verifies that the proposed method can correctly track locations of multiple speakers and meanwhile separate speech signals.

📄 PDF Abstract BibTeX arXiv:1710.10432

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TOGGL: Transcribing Overlapping Speech with Staggered Labeling

2024-08-12 · Chak-Fai Li, William Hartmann, Matthew Snover

Transcribing the speech of multiple overlapping speakers typically requires separating the audio into multiple streams and recognizing each one independently. More recent work jointly separates and transcribes, but requi…

AttributeDecoder

Multiple Speaker Separation from Noisy Sources in Reverberant Rooms using Relative Transfer Matrix

2025-03-12 · Wageesha N. Manamperi, Thushara D. Abhayapala

Separation of simultaneously active multiple speakers is a difficult task in environments with strong reverberation and many background noise sources. This paper uses the relative transfer matrix (ReTM), a generalization…

Speaker Separation

One-shot conditional audio filtering of arbitrary sounds

2020-11-04 · Beat Gfeller, Dominik Roblek, Marco Tagliasacchi

We consider the problem of separating a particular sound source from a single-channel mixture, based on only a short sample of the target source. Using SoundFilter, a wave-to-wave neural network architecture, we can trai…

MixCycle: Unsupervised Speech Separation via Cyclic Mixture Permutation Invariant Training

2022-02-08 · Ertuğ Karamatlı, Serap Kırbız

We introduce two unsupervised source separation methods, which involve self-supervised training from single-channel two-source speech mixtures. Our first method, mixture permutation invariant training (MixPIT), enables l…

Data AugmentationSpeech Separation

U-Mamba-Net: A highly efficient Mamba-based U-net style network for noisy and reverberant speech separation

2024-12-24 · Shaoxiang Dang, Tetsuya Matsumoto, Yoshinori Takeuchi, Hiroaki Kudo

The topic of speech separation involves separating mixed speech with multiple overlapping speakers into several streams, with each stream containing speech from only one speaker. Many highly effective models have emerged…

feature selectionMambaSpeech Separation