CPJD Corpus: Crowdsourced Parallel Speech Corpus of Japanese Dialects
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationSpeech RecognitionSpeech SynthesisSimilar Papers 제목 키워드 기반
A Crowdsourced Open-Source Kazakh Speech Corpus and Initial Speech Recognition Baseline
We present an open-source speech corpus for the Kazakh language. The Kazakh speech corpus (KSC) contains around 332 hours of transcribed audio comprising over 153,000 utterances spoken by participants from different regi…
speech-recognitionSpeech RecognitionFLEURS-R: A Restored Multilingual Speech Corpus for Generation Tasks
This paper introduces FLEURS-R, a speech restoration applied version of the Few-shot Learning Evaluation of Universal Representations of Speech (FLEURS) corpus. FLEURS-R maintains an N-way parallel speech corpus in 102 l…
Few-Shot Learningtext-to-speechText to SpeechDiscoGeM: A Crowdsourced Corpus of Genre-Mixed Implicit Discourse Relations
We present DiscoGeM, a crowdsourced corpus of 6,505 implicit discourse relations from three genres: political speech, literature, and encyclopedic texts. Each instance was annotated by 10 crowd workers. Various label agg…
RelationRelation ClassificationSpaceRef: A corpus of street-level geographic descriptions
This article describes SPACEREF, a corpus of street-level geographic descriptions. Pedestrians are walking a route in a (real) urban environment, describing their actions. Their position is automatically logged, their sp…
PositionAugmenting Librispeech with French Translations: A Multimodal Corpus for Direct Speech Translation Evaluation
Recent works in spoken language translation (SLT) have attempted to build end-to-end speech-to-text translation without using source language transcription during learning or decoding. However, while large quantities of …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSentence+5