paper-with-me

홈 › Papers

A large-scale and PCR-referenced vocal audio dataset for COVID-19

2022-12-15 · Jobie Budd, Kieran Baker, Emma Karoune, Harry Coppock, Selina Patel, Ana Tendero Cañadas, Alexander Titcomb, Richard Payne, David Hurley, Sabrina Egglestone, Lorraine Butler, Jonathon Mellor, George Nicholson, Ivan Kiskin, Vasiliki Koutra, Radka Jersakova, Rachel A. McKendry, Peter Diggle, Sylvia Richardson, Björn W. Schuller, Steven Gilmour, Davide Pigoli, Stephen Roberts, Josef Packham, Tracey Thornley, Chris Holmes

The UK COVID-19 Vocal Audio Dataset is designed for the training and evaluation of machine learning models that classify SARS-CoV-2 infection status or associated respiratory symptoms using vocal audio. The UK Health Security Agency recruited voluntary participants through the national Test and Trace programme and the REACT-1 survey in England from March 2021 to March 2022, during dominant transmission of the Alpha and Delta SARS-CoV-2 variants and some Omicron variant sublineages. Audio recordings of volitional coughs, exhalations, and speech were collected in the 'Speak up to help beat coronavirus' digital survey alongside demographic, self-reported symptom and respiratory condition data, and linked to SARS-CoV-2 test results. The UK COVID-19 Vocal Audio Dataset represents the largest collection of SARS-CoV-2 PCR-referenced audio recordings to date. PCR results were linked to 70,794 of 72,999 participants and 24,155 of 25,776 positive cases. Respiratory symptoms were reported by 45.62% of participants. This dataset has additional potential uses for bioacoustics research, with 11.30% participants reporting asthma, and 27.20% with linked influenza PCR test results.

📄 PDF Abstract BibTeX arXiv:2212.07738

Code (1)

alan-turing-institute/turing-rss-health-data-lab-biomedical-acoustic-markers 공식 구현 pytorch

Tasks

Survey

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Improving Query-by-Vocal Imitation with Contrastive Learning and Audio Pretraining

2024-08-21 · Jonathan Greif, Florian Schmid, Paul Primus, Gerhard Widmer

Query-by-Vocal Imitation (QBV) is about searching audio files within databases using vocal imitations created by the user's voice. Since most humans can effectively communicate sound concepts through voice, QBV offers th…

Contrastive Learning

Detection and classification of vocal productions in large scale audio recordings

2023-02-14 · Guillem Bonafos, Pierre Pudlo, Jean-Marc Freyermuth, Thierry Legou 외

We propose an automatic data processing pipeline to extract vocal productions from large-scale natural audio recordings and classify these vocal productions. The pipeline is based on a deep neural network and adresses bo…

Bayesian OptimisationData AugmentationTransfer Learning

Tadabur: A Large-Scale Quran Audio Dataset

2026-04-21 · Faisal Alherran arxiv

Despite growing interest in Quranic data research, existing Quran datasets remain limited in both scale and diversity. To address this gap, we present Tadabur, a large-scale Quran audio dataset. Tadabur comprises more th…

EMVD dataset: a dataset of extreme vocal distortion techniques used in heavy metal

2024-06-24 · Modan Tailleur, Julien Pinquier, Laurent Millot, Corsin Vogel 외

In this paper, we introduce the Extreme Metal Vocals Dataset, which comprises a collection of recordings of extreme vocal techniques performed within the realm of heavy metal music. The dataset consists of 760 audio exce…

VocalAgent: Large Language Models for Vocal Health Diagnostics with Safety-Aware Evaluation

2025-05-19 · Yubin Kim, Taehan Kim, Wonjune Kang, Eugene Park 외

Vocal health plays a crucial role in peoples' lives, significantly impacting their communicative abilities and interactions. However, despite the global prevalence of voice disorders, many lack access to convenient diagn…

DiagnosticLanguage ModelingLanguage ModellingLarge Language Model