paper-with-me

홈 › Papers

Adaptive blind audio source extraction supervised by dominant speaker identification using x-vectors

2019-10-25

We propose a novel algorithm for adaptive blind audio source extraction. The proposed method is based on independent vector analysis and utilizes the auxiliary function optimization to achieve high convergence speed. The algorithm is partially supervised by a pilot signal related to the source of interest (SOI), which ensures that the method correctly extracts the utterance of the desired speaker. The pilot is based on the identification of a dominant speaker in the mixture using x-vectors. The properties of the x-vectors computed in the presence of cross-talk are experimentally analyzed. The proposed approach is verified in a scenario with a moving SOI, static interfering speaker, and environmental noise.

📄 PDF Abstract BibTeX arXiv:1910.11824

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Identification

Similar Papers 제목 키워드 기반

Target Speech Extraction: Independent Vector Extraction Guided by Supervised Speaker Identification

2021-11-05 · Jiri Malek, Jakub Jansky, Zbynek Koldovsky, Tomas Kounovsky 외

This manuscript proposes a novel robust procedure for the extraction of a speaker of interest (SOI) from a mixture of audio sources. The estimation of the SOI is performed via independent vector extraction (IVE). Since t…

Speaker IdentificationSpeech Extraction

Neural Spectral Band Generation for Audio Coding

2025-06-07 · Woongjib Choi, Byeong Hyeon Kim, Hyungseob Lim, Inseon Jang 외

Audio bandwidth extension is the task of reconstructing missing high frequency components of bandwidth-limited audio signals, where bandwidth limitation is a common issue for audio signals due to several reasons, includi…

Bandwidth Extension

Generalized Canonical Correlation Analysis and Its Application to Blind Source Separation Based on a Dual-Linear Predictor Structure

2014-03-09 · Wei Liu

Blind source separation (BSS) is one of the most important and established research topics in signal processing and many algorithms have been proposed based on different statistical properties of the source signals. For …

blind source separation

ArrayDPS: Unsupervised Blind Speech Separation with a Diffusion Prior

2025-05-08 · Zhongweiyang Xu, Xulin Fan, Zhong-Qiu Wang, Xilin Jiang 외

Blind Speech Separation (BSS) aims to separate multiple speech sources from audio mixtures recorded by a microphone array. The problem is challenging because it is a blind inverse problem, i.e., the microphone array geom…

Room Impulse Response (RIR)Speech Separation

AudioSlots: A slot-centric generative model for audio separation

2023-05-09 · Pradyumna Reddy, Scott Wisdom, Klaus Greff, John R. Hershey 외

In a range of recent works, object-centric architectures have been shown to be suitable for unsupervised scene decomposition in the vision domain. Inspired by these methods we present AudioSlots, a slot-centric generativ…

blind source separationDecoderSpeech Separation