paper-with-me

홈 › Papers

Two-stage Neural Network for ICASSP 2023 Speech Signal Improvement Challenge

2023-03-14 · Mingshuai Liu, Shubo Lv, Zihan Zhang, Runduo Han, Xiang Hao, Xianjun Xia, Li Chen, Yijian Xiao, Lei Xie

In ICASSP 2023 speech signal improvement challenge, we developed a dual-stage neural model which improves speech signal quality induced by different distortions in a stage-wise divide-and-conquer fashion. Specifically, in the first stage, the speech improvement network focuses on recovering the missing components of the spectrum, while in the second stage, our model aims to further suppress noise, reverberation, and artifacts introduced by the first-stage model. Achieving 0.446 in the final score and 0.517 in the P.835 score, our system ranks 4th in the non-real-time track.

📄 PDF Abstract BibTeX arXiv:2303.07621

Code (0)

등록된 구현이 없습니다.

Tasks

Vocal Bursts Valence Prediction

Similar Papers 제목 키워드 기반

ICASSP 2023 Speech Signal Improvement Challenge

2023-03-12 · Ross Cutler, Ando Saabas, Babak Naderi, Nicolae-Cătălin Ristea 외

The ICASSP 2023 Speech Signal Improvement Challenge is intended to stimulate research in the area of improving the speech signal quality in communication systems. The speech signal quality can be measured with SIG in ITU…

ICASSP 2024 Speech Signal Improvement Challenge

2024-01-25 · Nicolae Catalin Ristea, Ando Saabas, Ross Cutler, Babak Naderi 외

The ICASSP 2024 Speech Signal Improvement Grand Challenge is intended to stimulate research in the area of improving the speech signal quality in communication systems. This marks our second challenge, building upon the …

BS-PLCNet 2: Two-stage Band-split Packet Loss Concealment Network with Intra-model Knowledge Distillation

2024-06-10 · Zihan Zhang, Xianjun Xia, Chuanzeng Huang, Yijian Xiao 외

Audio packet loss is an inevitable problem in real-time speech communication. A band-split packet loss concealment network (BS-PLCNet) targeting full-band signals was recently proposed. Although it performs superiorly in…

Knowledge DistillationPacket Loss Concealment

Speech Signal Improvement Using Causal Generative Diffusion Models

2023-03-15 · Julius Richter, Simon Welker, Jean-Marie Lemercier, Bunlong Lay 외

In this paper, we present a causal speech signal improvement system that is designed to handle different types of distortions. The method is based on a generative diffusion model which has been shown to work well in scen…

The PCG-AIID System for L3DAS22 Challenge: MIMO and MISO convolutional recurrent Network for Multi Channel Speech Enhancement and Speech Recognition

2022-02-21 · Jingdong Li, Yuanyuan Zhu, Dawei Luo, Yun Liu 외

This paper described the PCG-AIID system for L3DAS22 challenge in Task 1: 3D speech enhancement in office reverberant environment. We proposed a two-stage framework to address multi-channel speech denoising and dereverbe…

DenoisingSpeech DenoisingSpeech Enhancementspeech-recognition+1