paper-with-me

Papers

Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection

2025-03-10 · Michael Neri, Tuomas Virtanen

In this work, we investigate the generalization of a multi-channel learning-based replay speech detector, which employs adaptive beamforming and detection, across different microphone arrays. In general, deep neural network-based microphone array processing techniques generalize poorly to unseen array types, i.e., showing a significant training-test mismatch of performance. We employ the ReMASC dataset to analyze performance degradation due to inter- and intra-device mismatches, assessing both single- and multi-channel configurations. Furthermore, we explore fine-tuning to mitigate the performance loss when transitioning to unseen microphone arrays. Our findings reveal that array mismatches significantly decrease detection accuracy, with intra-device generalization being more robust than inter-device. However, fine-tuning with as little as ten minutes of target data can effectively recover performance, providing insights for practical deployment of replay detection systems in heterogeneous automatic speaker verification environments.

📄 PDF Abstract BibTeX arXiv:2503.07357

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Similar Papers 제목 키워드 기반

Libri-adhoc40: A dataset collected from synchronized ad-hoc microphone arrays

2021-03-28 · Shanzheng Guan, Shupei Liu, Junqi Chen, Wenbo Zhu 외

Recently, there is a research trend on ad-hoc microphone arrays. However, most research was conducted on simulated data. Although some data sets were collected with a small number of distributed devices, they were not sy…

speech-recognitionSpeech Recognition

SAMbA: Speech enhancement with Asynchronous ad-hoc Microphone Arrays

2023-07-31 · Nicolas Furnon, Romain Serizel, Slim Essid, Irina Illina

Speech enhancement in ad-hoc microphone arrays is often hindered by the asynchronization of the devices composing the microphone array. Asynchronization comes from sampling time offset and sampling rate offset which inev…

Speech Enhancement

Advances in Microphone Array Processing and Multichannel Speech Enhancement

2025-02-13 · Gongping Huang, Jesper R. Jensen, Jingdong Chen, Jacob Benesty 외

This paper reviews pioneering works in microphone array processing and multichannel speech enhancement, highlighting historical achievements, technological evolution, commercialization aspects, and key challenges. It pro…

Speech Enhancement

Attention-based distributed speech enhancement for unconstrained microphone arrays with varying number of nodes

2021-06-15 · Nicolas Furnon, Romain Serizel, Slim Essid, Irina Illina

Speech enhancement promises higher efficiency in ad-hoc microphone arrays than in constrained microphone arrays thanks to the wide spatial coverage of the devices in the acoustic scene. However, speech enhancement in ad-…

Speech Enhancement

Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer

2025-02-19 · Michael Neri, Tuomas Virtanen

Replay attacks belong to the class of severe threats against voice-controlled systems, exploiting the easy accessibility of speech signals by recorded and replayed speech to grant unauthorized access to sensitive data. I…