paper-with-me

VocalSound

홈페이지 · 논문 24편

VocalSound is a free dataset consisting of 21,024 crowdsourced recordings of laughter, sighs, coughs, throat clearing, sneezes, and sniffs from 3,365 unique subjects. The VocalSound dataset also contains meta-information such as speaker age, gender, native language, country, and health condition.

Audio

벤치마크

Audio Classification on VocalSound 결과 4개