paper-with-me

Speech Separation 벤치마크

Speech Separation on Libri2Mix

10개 결과 · ⬇ CSV · JSON

SI-SDRi

13.2 15.45 17.7 19.95 22.2 2020-10 2026-09 Conv-Tasnet (Libri1Mix speech enhancement pre-trained) — 14.1 (2020-10-29) Conv-Tasnet (Libri1Mix speech enhancement multi-task) — 13.7 (2020-10-29) Conv-Tasnet — 13.2 (2020-10-29) TDANet Large — 17.4 (2022-09-30) TDANet — 16.9 (2022-09-30) Separate And Diffuse — 21.5 (2023-01-25) MossFormer2 (w speed perturb) — 22.2 (2023-12-19) MossFormer2 (w/o DM) — 21.7 (2023-12-19) TF-Locoformer (M) — 22.1 (2024-08-06) WHYV — 17.5 (2024-10-01) Conv-Tasnet (Libri1Mix speech enhancement pre-trained) — 14.1 (2020-10-29) TDANet Large — 17.4 (2022-09-30) Separate And Diffuse — 21.5 (2023-01-25) MossFormer2 (w speed perturb) — 22.2 (2023-12-19)
RankModel SI-SDRiSDRiNumber of parameters (M)SDR Extra Training Data PaperCodeYear
1 MossFormer2 (w speed perturb) 22.2 MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation modelscope/ClearerVoice-Studio · alibabasglab/MossFormer2 2023
2 TF-Locoformer (M) 22.122.215 TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement merlresearch/tf-locoformer 2024
3 MossFormer2 (w/o DM) 21.7 MossFormer2: Combining Transformer and RNN-Free Recurrent Network for Enhanced Time-Domain Monaural Speech Separation modelscope/ClearerVoice-Studio · alibabasglab/MossFormer2 2023
4 Separate And Diffuse 21.5 Separate And Diffuse: Using a Pretrained Diffusion Model for Improving Source Separation 2023
5 WHYV 17.517.2458 Wanna hear your voice? A sample is all we need! 2024
6 TDANet Large 17.4 An efficient encoder-decoder architecture with top-down attention for speech separation JusperLee/TDANet 2022
7 TDANet 16.9 An efficient encoder-decoder architecture with top-down attention for speech separation JusperLee/TDANet 2022
8 Conv-Tasnet (Libri1Mix speech enhancement pre-trained) 14.114.6 Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training SungFeng-Huang/SSL-pretraining-separation 2020
9 Conv-Tasnet (Libri1Mix speech enhancement multi-task) 13.714.1 Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training SungFeng-Huang/SSL-pretraining-separation 2020
10 Conv-Tasnet 13.213.6 Stabilizing Label Assignment for Speech Separation by Self-supervised Pre-training SungFeng-Huang/SSL-pretraining-separation 2020
1–10 / 10 페이지당 10 20 50 100