paper-with-me

Papers

Codec Data Augmentation for Time-domain Heart Sound Classification

2023-09-14 · Ansh Mishra, Jia Qi Yip, Eng Siong Chng

Heart auscultations are a low-cost and effective way of detecting valvular heart diseases early, which can save lives. Nevertheless, it has been difficult to scale this screening method since the effectiveness of auscultations is dependent on the skill of doctors. As such, there has been increasing research interest in the automatic classification of heart sounds using deep learning algorithms. However, it is currently difficult to develop good heart sound classification models due to the limited data available for training. In this work, we propose a simple time domain approach, to the heart sound classification problem with a base classification error rate of 0.8 and show that augmentation of the data through codec simulation can improve the classification error rate to 0.2. With data augmentation, our approach outperforms the existing time-domain CNN-BiLSTM baseline model. Critically, our experiments show that codec data augmentation is effective in getting around the data limitation.

📄 PDF Abstract BibTeX arXiv:2309.07466

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationData AugmentationSound Classification

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Towards Fusion of Neural Audio Codec-based Representations with Spectral for Heart Murmur Classification via Bandit-based Cross-Attention Mechanism

2025-06-01 · Orchid Chetia Phukan, Girish, Mohd Mujtaba Akhtar, Swarup Ranjan Behera 외

In this study, we focus on heart murmur classification (HMC) and hypothesize that combining neural audio codec representations (NACRs) such as EnCodec with spectral features (SFs), such as MFCC, will yield superior perfo…

Rhythm

Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis

2024-09-20 · Lauri Juvela, Xin Wang

Automatic detection of synthetic speech is becoming increasingly important as current synthesis methods are both near indistinguishable from human speech and widely accessible to the public. Audio watermarking and other …

Face SwappingSpeech Synthesis

A Residual Diffusion Model for High Perceptual Quality Codec Augmentation

2023-01-13 · Noor Fathima Ghouse, Jens Petersen, Auke Wiggers, Tianlin Xu 외

Diffusion probabilistic models have recently achieved remarkable success in generating high quality image and video data. In this work, we build on this class of generative models and introduce a method for lossy compres…

Image CompressionVocal Bursts Intensity Prediction

Augmentation-based Domain Generalization and Joint Training from Multiple Source Domains for Whole Heart Segmentation

2025-08-06 · Franz Thaler, Darko Stern, Gernot Plank, Martin Urschler arxiv

As the leading cause of death worldwide, cardiovascular diseases motivate the development of more sophisticated methods to analyze the heart and its substructures from medical images like Computed Tomography (CT) and Mag…

Medical Image SegmentationDomain Generalization

Cardiac Disease Diagnosis on Imbalanced Electrocardiography Data Through Optimal Transport Augmentation

2022-01-25 · JieLin Qiu, Jiacheng Zhu, Mengdi Xu, Peide Huang 외

In this paper, we focus on a new method of data augmentation to solve the data imbalance problem within imbalanced ECG datasets to improve the robustness and accuracy of heart disease detection. By using Optimal Transpor…

Data Augmentation