A Crowdsourcing-based Approach for Speech Corpus Transcription Case of Arabic Algerian Dialects
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Best Practices for Crowdsourcing Dialectal Arabic Speech Transcription
Can Crowdsourcing be used for Effective Annotation of Arabic?
Crowdsourcing has been used recently as an alternative to traditional costly annotation by many natural language processing groups. In this paper, we explore the use of Amazon Mechanical Turk (AMT) in order to assess the…
Entity ResolutionMultiple-choiceNatural Language InferencePart-Of-Speech Tagging+3ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
We introduce ArVoice, a multi-speaker Modern Standard Arabic (MSA) speech corpus with diacritized transcriptions, intended for multi-speaker speech synthesis, and can be useful for other tasks such as speech-based diacri…
DeepFake DetectionFace SwappingSpeech SynthesisVoice ConversionQASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus
We introduce the largest transcribed Arabic speech corpus, QASR, collected from the broadcast domain. This multi-dialect speech dataset contains 2,000 hours of speech sampled at 16kHz crawled from Aljazeera news channel.…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationLanguage Modeling+8QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus
We introduce the largest transcribed Arabic speech corpus, QASR, collected from the broadcast domain. This multi-dialect speech dataset contains 2,000 hours of speech sampled at 16kHz crawled from Aljazeera news channel.…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationLanguage Modeling+8