paper-with-me

홈 › Papers

LMFCA-Net: A Lightweight Model for Multi-Channel Speech Enhancement with Efficient Narrow-Band and Cross-Band Attention

2025-02-17 · Yaokai Zhang, Hanchen Pei, Wanqi Wang, Gongping Huang

Deep learning based end-to-end multi-channel speech enhancement methods have achieved impressive performance by leveraging sub-band, cross-band, and spatial information. However, these methods often demand substantial computational resources, limiting their practicality on terminal devices. This paper presents a lightweight multi-channel speech enhancement network with decoupled fully connected attention (LMFCA-Net). The proposed LMFCA-Net introduces time-axis decoupled fully-connected attention (T-FCA) and frequency-axis decoupled fully-connected attention (F-FCA) mechanisms to effectively capture long-range narrow-band and cross-band information without recurrent units. Experimental results show that LMFCA-Net performs comparably to state-of-the-art methods while significantly reducing computational complexity and latency, making it a promising solution for practical applications.

📄 PDF Abstract BibTeX arXiv:2502.11462

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Enhancement

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

A Lightweight Hybrid Dual Channel Speech Enhancement System under Low-SNR Conditions

2025-05-26 · Zheng Wang, Xiaobin Rong, Yu Sun, Tianchi Sun 외

Although deep learning based multi-channel speech enhancement has achieved significant advancements, its practical deployment is often limited by constrained computational resources, particularly in low signal-to-noise r…

Speech Enhancement

Incorporating Multi-Target in Multi-Stage Speech Enhancement Model for Better Generalization

2021-07-09 · Lu Zhang, Mingjiang Wang, Andong Li, Zehua Zhang 외

Recent single-channel speech enhancement methods based on deep neural networks (DNNs) have achieved remarkable results, but there are still generalization problems in real scenes. Like other data-driven methods, DNN-base…

DenoisingSpeech DenoisingSpeech Enhancement

FullSubNet+: Channel Attention FullSubNet with Complex Spectrograms for Speech Enhancement

2022-03-23 · Jun Chen, Zilin Wang, Deyi Tuo, Zhiyong Wu 외

Previously proposed FullSubNet has achieved outstanding performance in Deep Noise Suppression (DNS) Challenge and attracted much attention. However, it still encounters issues such as input-output mismatch and coarse pro…

Speech Enhancement

Student-Teacher Learning for BLSTM Mask-based Speech Enhancement

2018-03-27

Spectral mask estimation using bidirectional long short-term memory (BLSTM) neural networks has been widely used in various speech enhancement applications, and it has achieved great success when it is applied to multich…

Speech Enhancementspeech-recognitionSpeech Recognition

Closing the Gap Between Time-Domain Multi-Channel Speech Enhancement on Real and Simulation Conditions

2021-10-27 · Wangyou Zhang, Jing Shi, Chenda Li, Shinji Watanabe 외

The deep learning based time-domain models, e.g. Conv-TasNet, have shown great potential in both single-channel and multi-channel speech enhancement. However, many experiments on the time-domain speech enhancement model …

Speech Enhancementspeech-recognitionSpeech Recognition