paper-with-me

홈 › Papers

NASTAR: Noise Adaptive Speech Enhancement with Target-Conditional Resampling

2022-06-18 · Chi-Chang Lee, Cheng-Hung Hu, Yu-Chen Lin, Chu-Song Chen, Hsin-Min Wang, Yu Tsao

For deep learning-based speech enhancement (SE) systems, the training-test acoustic mismatch can cause notable performance degradation. To address the mismatch issue, numerous noise adaptation strategies have been derived. In this paper, we propose a novel method, called noise adaptive speech enhancement with target-conditional resampling (NASTAR), which reduces mismatches with only one sample (one-shot) of noisy speech in the target environment. NASTAR uses a feedback mechanism to simulate adaptive training data via a noise extractor and a retrieval model. The noise extractor estimates the target noise from the noisy speech, called pseudo-noise. The noise retrieval model retrieves relevant noise samples from a pool of noise signals according to the noisy speech, called relevant-cohort. The pseudo-noise and the relevant-cohort set are jointly sampled and mixed with the source speech corpus to prepare simulated training data for noise adaptation. Experimental results show that NASTAR can effectively use one noisy speech sample to adapt an SE model to a target condition. Moreover, both the noise extractor and the noise retrieval model contribute to model adaptation. To our best knowledge, NASTAR is the first work to perform one-shot noise adaptation through noise extraction and retrieval.

📄 PDF Abstract BibTeX arXiv:2206.09058

Code (0)

등록된 구현이 없습니다.

Tasks

RetrievalSpeech Enhancement

Similar Papers 제목 키워드 기반

Effective Noise-aware Data Simulation for Domain-adaptive Speech Enhancement Leveraging Dynamic Stochastic Perturbation

2024-09-03 · Chien-Chun Wang, Li-Wei Chen, Hung-Shin Lee, Berlin Chen 외

Cross-domain speech enhancement (SE) is often faced with severe challenges due to the scarcity of noise and background information in an unseen target domain, leading to a mismatch between training and test conditions. T…

Speech Enhancement

Blind Mask to Improve Intelligibility of Non-Stationary Noisy Speech

2020-08-20 · F. Farias, R. Coelho

This letter proposes a novel blind acoustic mask (BAM) designed to adaptively detect noise components and preserve target speech segments in time-domain. A robust standard deviation estimator is applied to the non-statio…

Noise EstimationSpeech Enhancement

Streaming Noise Context Aware Enhancement For Automatic Speech Recognition in Multi-Talker Environments

2022-05-17 · Joe Caroselli, Arun Narayanan, Yiteng Huang

One of the most challenging scenarios for smart speakers is multi-talker, when target speech from the desired speaker is mixed with interfering speech from one or more speakers. A smart assistant needs to determine which…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+1

Speech Enhancement using Adaptive Mean Median Deviation and EMD Technique

2019-08-26 · Conference 2019 8 · Shikha Dubey, Ashish Kumar Singh, Manoj Kumar Singh

During the acquisition of the speech signal by the non-contact Speech Sensor (SS), the signal is degraded by severe colored noises which are non-linear and non-uniform in nature. Therefore, in this study, a new approach …

Speech Enhancement

NASTaR: NovaSAR Automated Ship Target Recognition Dataset

2025-12-20 · Benyamin Hosseiny, Kamirul Kamirul, Odysseas Pappas, Alin Achim arxiv

Synthetic Aperture Radar (SAR) offers a unique capability for all-weather, space-based maritime activity monitoring by capturing and imaging strong reflections from ships at sea. A well-defined challenge in this domain i…