paper-with-me

Papers

Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases

2024-05-13 · Pengfei Zhang, Zhihang Zheng, Shichen Zhang, Minghao Yang, Shaojun Tang

Compared with invasive examinations that require tissue sampling, respiratory sound testing is a non-invasive examination method that is safer and easier for patients to accept. In this study, we introduce Rene, a pioneering large-scale model tailored for respiratory sound recognition. Rene has been rigorously fine-tuned with an extensive dataset featuring a broad array of respiratory audio samples, targeting disease detection, sound pattern classification, and event identification. Our innovative approach applies a pre-trained speech recognition model to process respiratory sounds, augmented with patient medical records. The resulting multi-modal deep-learning framework addresses interpretability and real-time diagnostic challenges that have hindered previous respiratory-focused models. Benchmark comparisons reveal that Rene significantly outperforms existing models, achieving improvements of 10.27%, 16.15%, 15.29%, and 18.90% in respiratory event detection and audio classification on the SPRSound database. Disease prediction accuracy on the ICBHI database improved by 23% over the baseline in both mean average and harmonic scores. Moreover, we have developed a real-time respiratory sound discrimination system utilizing the Rene architecture. Employing state-of-the-art Edge AI technology, this system enables rapid and accurate responses for respiratory sound auscultation(https://github.com/zpforlove/Rene).

📄 PDF Abstract BibTeX arXiv:2405.07442

Code (1)

zpforlove/rene 공식 구현 pytorch

Tasks

Audio ClassificationDiagnosticDisease PredictionEvent Detectionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Patient-Level Multimodal Question Answering from Multi-Site Auscultation Recordings

2026-03-09 · Fan Wu, Tsai-Ning Wang, Nicolas Zumarraga, Ning Wang 외 arxiv

Auscultation is a vital diagnostic tool, yet its utility is often limited by subjective interpretation. While general-purpose Audio-Language Models (ALMs) excel in general domains, they struggle with the nuances of physi…

Question Answering

Interactive Lungs Auscultation with Reinforcement Learning Agent

2019-07-25 · Tomasz Grzywalski, Riccardo Belluzzo, Szymon Drgas, Agnieszka Cwalinska 외

To perform a precise auscultation for the purposes of examination of respiratory system normally requires the presence of an experienced doctor. With most recent advances in machine learning and artificial intelligence, …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

StethoLM: Audio Language Model for Cardiopulmonary Analysis Across Clinical Tasks

2026-02-27 · Yishan Wang, Tsai-Ning Wang, Mathias Funk, Aaqib Saeed arxiv

Listening to heart and lung sounds - auscultation - is one of the first and most fundamental steps in a clinical examination. Despite being fast and non-invasive, it demands years of experience to interpret subtle audio …

Binary Classification

Scattering Transformer: A Training-Free Transformer Architecture for Heart Murmur Detection

2025-09-22 · Rami Zewail arxiv

In an attempt to address the need for skilled clinicians in heart sound interpretation, recent research efforts on automating cardiac auscultation have explored deep learning approaches. The majority of these approaches …

Deep auscultation: Predicting respiratory anomalies and diseases via recurrent neural networks

2019-07-11 · Diego Perna, Andrea Tagarelli

Respiratory diseases are among the most common causes of severe illness and death worldwide. Prevention and early diagnosis are essential to limit or even reverse the trend that characterizes the diffusion of such diseas…