paper-with-me

VibraVox (rigid in-ear microphone)

홈페이지 · 논문 1편

This is the in-ear rigid earpiece-embedded microphone variant of the VibraVox dataset. VibraVox aims at serving as a valuable resource for advancing the field of body-conducted speech analysis and facilitating the development of robust communication systems for real-world applications. This dataset, available on HuggingFace can be used for various audio machine learning tasks : - Automatic Speech Recognition (ASR) (Speech-to-Text , Speech-to-Phoneme) - Audio Bandwidth Extension (BWE) - Speaker Verification (SPKV) / identification - Voice cloning - etc ... The VibraVox dataset speech corpus has been released in July 2024. It includes french speech recorded simultaneously using multiple audio and vibration sensors : a forehead miniature vibration sensor, an in-ear comply foam-embedded microphone, an in-ear rigid earpiece-embedded microphone, a temple vibration pickup, a headset microphone located near the mouth, and a laryngophone. VibraVox has been recorded with 200 participants under various acoustic conditions imposed by a 5th order ambisonics spatialization sphere.

TextsAudioSpeech French

벤치마크

Automatic Phoneme Recognition on VibraVox (rigid in-ear microphone) 결과 2개
Bandwidth Extension on VibraVox (rigid in-ear microphone) 결과 1개
Speaker Verification on VibraVox (rigid in-ear microphone) 결과 1개