Speaker and Posture Classification using Instantaneous Intraspeech Breathing Features
Acoustic features extracted from speech are widely used in problems such as biometric speaker identification and first-person activity detection. However, the use of speech for such purposes raises privacy issues as the content is accessible to the processing party. In this work, we propose a method for speaker and posture classification using intraspeech breathing sounds. Instantaneous magnitude features are extracted using the Hilbert-Huang transform (HHT) and fed into a CNN-GRU network for classification of recordings from the open intraspeech breathing sound dataset, BreathBase, that we collected for this study. Using intraspeech breathing sounds, 87% speaker classification, and 98% posture classification accuracy were obtained.
Code (0)
등록된 구현이 없습니다.
Tasks
Action DetectionActivity DetectionClassificationGeneral ClassificationSpeaker IdentificationSimilar Papers 제목 키워드 기반
Remote Breathing Monitoring Using LiDAR Technology
Breathing monitoring is crucial in healthcare for early detection of health issues, but traditional methods face challenges like invasiveness, privacy concerns, and limited applicability in daily settings. This paper int…
PositionMultitask Network for Respiration Rate Estimation -- A Practical Perspective
The exponential rise in wearable sensors has garnered significant interest in assessing the physiological parameters during day-to-day activities. Respiration rate is one of the vital parameters used in the performance a…
DecoderDetecting COVID-19 from Breathing and Coughing Sounds using Deep Neural Networks
The COVID-19 pandemic has affected the world unevenly; while industrial economies have been able to produce the tests necessary to track the spread of the virus and mostly avoided complete lockdowns, developing countries…
Bayesian OptimisationDynamically locating multiple speakers based on the time-frequency domain
In this study we present a deep neural network-based online multi-speaker localisation algorithm based on a multi-microphone array. A fully convolutional network is trained with instantaneous spatial features to estimate…
FCN Approach for Dynamically Locating Multiple Speakers
In this paper, we present a deep neural network-based online multi-speaker localisation algorithm. Following the W-disjoint orthogonality principle in the spectral domain, each time-frequency (TF) bin is dominated by a s…