VibraVox (forehead accelerometer)
홈페이지 · 논문 1편
This is the forehead accelerometer variant of the VibraVox dataset. VibraVox aims at serving as a valuable resource for advancing the field of body-conducted speech analysis and facilitating the development of robust communication systems for real-world applications. This dataset, available on HuggingFace can be used for various audio machine learning tasks : - Automatic Speech Recognition (ASR) (Speech-to-Text , Speech-to-Phoneme) - Audio Bandwidth Extension (BWE) - Speaker Verification (SPKV) / identification - Voice cloning - etc ... The VibraVox dataset speech corpus has been released in July 2024. It includes french speech recorded simultaneously using multiple audio and vibration sensors : a forehead miniature vibration sensor, an in-ear comply foam-embedded microphone, an in-ear rigid earpiece-embedded microphone, a temple vibration pickup, a headset microphone located near the mouth, and a laryngophone. VibraVox has been recorded with 200 participants under various acoustic conditions imposed by a 5th order ambisonics spatialization sphere.
TextsAudioSpeech French