Improving the Robustness and Clinical Applicability of Automatic Respiratory Sound Classification Using Deep Learning-Based Audio Enhancement: Algorithm Development and Validation
Deep learning techniques have shown promising results in the automatic classification of respiratory sounds. However, accurately distinguishing these sounds in real-world noisy conditions remains challenging for clinical deployment. In addition, predicting signals with only background noise may reduce user trust in the system. This study explores the feasibility and effectiveness of incorporating a deep learning-based audio enhancement step into automatic respiratory sound classification systems to improve robustness and clinical applicability. We conducted extensive experiments using various audio enhancement model architectures, including time-domain and time-frequency-domain approaches, combined with multiple classification models to evaluate the module's effectiveness. The classification performance was compared against the noise injection data augmentation method. These experiments were carried out on two datasets: the ICBHI respiratory sound dataset and the FABS dataset. Furthermore, a physician validation study assessed the system's clinical utility. Integrating the audio enhancement module resulted in a 21.9% increase in the ICBHI classification score and a 4.1% improvement on the FABS dataset in multi-class noisy scenarios. Quantitative analysis revealed efficiency gains, higher diagnostic confidence, and increased trust, with workflows using enhanced audio improving diagnostic sensitivity by 11.6% and enabling high-confidence diagnoses. Incorporating an audio enhancement algorithm boosts the robustness and clinical utility of automatic respiratory sound classification systems, enhancing performance in noisy environments and fostering greater trust among medical professionals.
Code (0)
등록된 구현이 없습니다.
Tasks
ClassificationData AugmentationDiagnosticSound ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
EZhouNet:A framework based on graph neural network and anchor interval for the respiratory sound event detection
Auscultation is a key method for early diagnosis of respiratory and pulmonary diseases, relying on skilled healthcare professionals. However, the process is often subjective, with variability between experts. As a result…
Sound Event DetectionGraph Neural NetworkAutomatic Classification of Large-Scale Respiratory Sound Dataset Based on Convolutional Neural Network
Auscultation of respiratory sounds is very important for discovering the respiratory disease. However, there is no quantitative evaluation method for the diagnosis of respiratory sounds until now. It is necessary to deve…
Audio ClassificationClassificationSpecificityTowards Enhanced Classification of Abnormal Lung sound in Multi-breath: A Light Weight Multi-label and Multi-head Attention Classification Method
This study aims to develop an auxiliary diagnostic system for classifying abnormal lung respiratory sounds, enhancing the accuracy of automatic abnormal breath sound classification through an innovative multi-label learn…
DiagnosticDiversityMulti-Label LearningSound ClassificationSPRSound: Open-Source SJTU Paediatric Respiratory Sound Database
It has proved that the auscultation of respiratory sound has advantage in early respiratory diagnosis. Various methods have been raised to perform automatic respiratory sound analysis to reduce subjective diagnosis and p…
ClassificationSound ClassificationMeta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification
Training reliable respiratory sound classification models remains challenging due to the limited size and subject diversity of datasets. Ensemble methods can improve robustness, but when base models are trained on identi…
Ensemble Learning