paper-with-me

Papers

A Crowdsourcing-based Approach for Speech Corpus Transcription Case of Arabic Algerian Dialects

2019-09-01 · WS 2019 9 · Ilyes Zine, Mohamed Cherif Zeghad, Soumia Bougrine, Hadda Cherroun
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Best Practices for Crowdsourcing Dialectal Arabic Speech Transcription

2015-07-01 · WS 2015 7 · Samantha Wray, Hamdy Mubarak, Ahmed Ali
Speech Recognition

Can Crowdsourcing be used for Effective Annotation of Arabic?

2014-05-01 · LREC 2014 5 · Wajdi Zaghouani, Kais Dukes

Crowdsourcing has been used recently as an alternative to traditional costly annotation by many natural language processing groups. In this paper, we explore the use of Amazon Mechanical Turk (AMT) in order to assess the…

Entity ResolutionMultiple-choiceNatural Language InferencePart-Of-Speech Tagging+3

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

2025-05-26 · Hawau Olamide Toyin, Rufael Marew, Humaid Alblooshi, Samar M. Magdy 외

We introduce ArVoice, a multi-speaker Modern Standard Arabic (MSA) speech corpus with diacritized transcriptions, intended for multi-speaker speech synthesis, and can be useful for other tasks such as speech-based diacri…

DeepFake DetectionFace SwappingSpeech SynthesisVoice Conversion

QASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus

2021-08-01 · ACL 2021 5 · Hamdy Mubarak, Amir Hussein, Shammur Absar Chowdhury, Ahmed Ali

We introduce the largest transcribed Arabic speech corpus, QASR, collected from the broadcast domain. This multi-dialect speech dataset contains 2,000 hours of speech sampled at 16kHz crawled from Aljazeera news channel.…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationLanguage Modeling+8

QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus

2021-06-24 · Hamdy Mubarak, Amir Hussein, Shammur Absar Chowdhury, Ahmed Ali

We introduce the largest transcribed Arabic speech corpus, QASR, collected from the broadcast domain. This multi-dialect speech dataset contains 2,000 hours of speech sampled at 16kHz crawled from Aljazeera news channel.…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationLanguage Modeling+8