paper-with-me

LibriSpeech

홈페이지 · 논문 2,361편

The LibriSpeech corpus is a collection of approximately 1,000 hours of audiobooks that are a part of the LibriVox project. Most of the audiobooks come from the Project Gutenberg. The training data is split into 3 partitions of 100hr, 360hr, and 500hr sets while the dev and test data are split into the ’clean’ and ’other’ categories, respectively, depending upon how well or challenging Automatic Speech Recognition systems would perform against. Each of the dev and test sets is around 5hr in audio length. This corpus also provides the n-gram language models and the corresponding texts excerpted from the Project Gutenberg books, which contain 803M tokens and 977K unique words. Source: State-of-the-art Speech Recognition using Multi-stream Self-attention with Dilated 1D Convolutions

AudioSpeech EnglishFrenchSpanishItalianJapanesePortuguese

벤치마크

Speech Recognition on LibriSpeech test-clean 결과 66개
Speech Recognition on LibriSpeech test-other 결과 53개
Voice Conversion on LibriSpeech test-clean 결과 3개