paper-with-me

Papers

Interspeech 2025 URGENT Speech Enhancement Challenge

2025-05-29 · Kohei Saijo, Wangyou Zhang, Samuele Cornell, Robin Scheibler, Chenda Li, Zhaoheng Ni, Anurag Kumar, Marvin Sach, Yihui Fu, Wei Wang, Tim Fingscheidt, Shinji Watanabe

There has been a growing effort to develop universal speech enhancement (SE) to handle inputs with various speech distortions and recording conditions. The URGENT Challenge series aims to foster such universal SE by embracing a broad range of distortion types, increasing data diversity, and incorporating extensive evaluation metrics. This work introduces the Interspeech 2025 URGENT Challenge, the second edition of the series, to explore several aspects that have received limited attention so far: language dependency, universality for more distortion types, data scalability, and the effectiveness of using noisy training data. We received 32 submissions, where the best system uses a discriminative model, while most other competitive ones are hybrid methods. Analysis reveals some key findings: (i) some generative or hybrid approaches are preferred in subjective evaluations over the top discriminative model, and (ii) purely generative SE models can exhibit language dependency.

📄 PDF Abstract BibTeX arXiv:2505.23212

Code (0)

등록된 구현이 없습니다.

Tasks

DiversitySpeech Enhancement

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

P.808 Multilingual Speech Enhancement Testing: Approach and Results of URGENT 2025 Challenge

2025-07-15 · Marvin Sach, Yihui Fu, Kohei Saijo, Wangyou Zhang 외

In speech quality estimation for speech enhancement (SE) systems, subjective listening tests so far are considered as the gold standard. This should be even more true considering the large influx of new generative or hyb…

Speech Enhancementtext-to-speechText to Speech

TS-URGENet: A Three-stage Universal Robust and Generalizable Speech Enhancement Network

2025-05-24 · Xiaobin Rong, DaHan Wang, Qinwen Hu, Yushi Wang 외

Universal speech enhancement aims to handle input speech with different distortions and input formats. To tackle this challenge, we present TS-URGENet, a Three-Stage Universal, Robust, and Generalizable speech Enhancemen…

Speech Enhancement

Challenges and Opportunities in Multi-device Speech Processing

2022-06-27 · Gregory Ciccarelli, Jarred Barber, Arun Nair, Israel Cohen 외

We review current solutions and technical challenges for automatic speech recognition, keyword spotting, device arbitration, speech enhancement, and source localization in multidevice home environments to provide context…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Keyword SpottingSpeech Enhancement+2

FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement

2020-10-29 · Xiang Hao, Xiangdong Su, Radu Horaud, Xiaofei Li

This paper proposes a full-band and sub-band fusion model, named as FullSubNet, for single-channel real-time speech enhancement. Full-band and sub-band refer to the models that input full-band and sub-band noisy spectral…

Speech Enhancement

The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Speech Quality and Testing Framework

2020-01-23 · Chandan K. A. Reddy, Ebrahim Beyrami, Harishchandra Dubey, Vishak Gopal 외

The INTERSPEECH 2020 Deep Noise Suppression Challenge is intended to promote collaborative research in real-time single-channel Speech Enhancement aimed to maximize the subjective (perceptual) quality of the enhanced spe…

Speech Enhancement