paper-with-me

홈 › Papers

TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition

2024-04-19 · Chengxin Chen, Pengyuan Zhang

One persistent challenge in Speech Emotion Recognition (SER) is the ubiquitous environmental noise, which frequently results in deteriorating SER performance in practice. In this paper, we introduce a Two-level Refinement Network, dubbed TRNet, to address this challenge. Specifically, a pre-trained speech enhancement module is employed for front-end noise reduction and noise level estimation. Later, we utilize clean speech spectrograms and their corresponding deep representations as reference signals to refine the spectrogram distortion and representation shift of enhanced speech during model training. Experimental results validate that the proposed TRNet substantially promotes the robustness of the proposed system in both matched and unmatched noisy environments, without compromising its performance in noise-free environments.

📄 PDF Abstract BibTeX arXiv:2404.12979

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionSpeech Emotion RecognitionSpeech Enhancement

Similar Papers 제목 키워드 기반

Cross-Talk Reduction

2024-05-30 · Zhong-Qiu Wang, Anurag Kumar, Shinji Watanabe

While far-field multi-talker mixtures are recorded, each speaker can wear a close-talk microphone so that close-talk mixtures can be recorded at the same time. Although each close-talk mixture has a high signal-to-noise …

Speech Separation

PGTRNet: Two-phase Weakly Supervised Object Detection with Pseudo Ground Truth Refinement

2021-08-25 · Jun Wang, Hefeng Zhou, Xiaohan Yu

Current state-of-the-art weakly supervised object detection (WSOD) studies mainly follow a two-stage training strategy which integrates a fully supervised detector (FSD) with a pure WSOD model. There are two main problem…

object-detectionObject DetectionWeakly Supervised Object Detection

Heterogeneous Space Fusion and Dual-Dimension Attention: A New Paradigm for Speech Enhancement

2024-08-13 · Tao Zheng, Liejun Wang, Yinfeng Yu

Self-supervised learning has demonstrated impressive performance in speech tasks, yet there remains ample opportunity for advancement in the realm of speech enhancement research. In addressing speech tasks, confining the…

Self-Supervised LearningSpeech Enhancement

DTRNet: Dual Text-Radical Decoding for Handwritten Chinese Text Recognition with Faked Character Detection

2026-08-06 · Runrui Li, Lin Zhu, Hua Huang arxiv

In K-12 educational scenarios, handwritten Chinese text recognition should not only transcribe student writing, but also detect faked characters. However, existing recognition models are usually confined to a predefined …

Joint Anchor-Feature Refinement for Real-Time Accurate Object Detection in Images and Videos

2018-07-23 · Xingyu Chen, Junzhi Yu, Shihan Kong, Zhengxing Wu 외

Object detection has been vigorously investigated for years but fast accurate detection for real-world scenes remains a very challenging problem. Overcoming drawbacks of single-stage detectors, we take aim at precisely d…

Objectobject-detectionObject DetectionReal-Time Object Detection