paper-with-me

VibraVox (throat microphone)

홈페이지 · 논문 1편

This is the throat microphone (laryngophone) variant of the VibraVox dataset. VibraVox aims at serving as a valuable resource for advancing the field of body-conducted speech analysis and facilitating the development of robust communication systems for real-world applications. This dataset, available on HuggingFace can be used for various audio machine learning tasks : - Automatic Speech Recognition (ASR) (Speech-to-Text , Speech-to-Phoneme) - Audio Bandwidth Extension (BWE) - Speaker Verification (SPKV) / identification - Voice cloning - etc ... The VibraVox dataset speech corpus has been released in July 2024. It includes french speech recorded simultaneously using multiple audio and vibration sensors : a forehead miniature vibration sensor, an in-ear comply foam-embedded microphone, an in-ear rigid earpiece-embedded microphone, a temple vibration pickup, a headset microphone located near the mouth, and a laryngophone. VibraVox has been recorded with 200 participants under various acoustic conditions imposed by a 5th order ambisonics spatialization sphere.

TextsAudioSpeech French

벤치마크

Automatic Phoneme Recognition on VibraVox (throat microphone) 결과 2개
Bandwidth Extension on VibraVox (throat microphone) 결과 1개
Speaker Verification on VibraVox (throat microphone) 결과 1개