paper-with-me

Papers

A study for the effect of the Emphaticness and language and dialect for Voice Onset Time (VOT) in Modern Standard Arabic (MSA)

2013-05-13 · Sulaiman S. AlDahri

The signal sound contains many different features, including Voice Onset Time (VOT), which is a very important feature of stop sounds in many languages. The only application of VOT values is stopping phoneme subsets. This subset of consonant sounds is stop phonemes exist in the Arabic language, and in fact, all languages. The pronunciation of these sounds is hard and unique especially for less-educated Arabs and non-native Arabic speakers. VOT can be utilized by the human auditory system to distinguish between voiced and unvoiced stops such as /p/ and /b/ in English.This search focuses on computing and analyzing VOT of Modern Standard Arabic (MSA), within the Arabic language, for all pairs of non-emphatic (namely, /d/ and /t/) and emphatic pairs (namely, /d?/ and /t?/) depending on carrier words. This research uses a database built by ourselves, and uses the carrier words syllable structure: CV-CV-CV. One of the main outcomes always found is the emphatic sounds (/d?/, /t?/) are less than 50% of non-emphatic (counter-part) sounds ( /d/, /t/).Also, VOT can be used to classify or detect for a dialect ina language.

📄 PDF Abstract BibTeX arXiv:1305.2680

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Crowdsourcing Latin American Spanish for Low-Resource Text-to-Speech

2020-05-01 · LREC 2020 5 · Adriana Guevara-Rukoz, Isin Demirsahin, Fei He, Shan-Hui Cathy Chu 외

In this paper we present a multidialectal corpus approach for building a text-to-speech voice for a new dialect in a language with existing resources, focusing on various South American dialects of Spanish. We first pres…

text-to-speechText to Speech

Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT

2024-09-11 · Kazuki Yamauchi, Yuki Saito, Hiroshi Saruwatari

We explore cross-dialect text-to-speech (CD-TTS), a task to synthesize learned speakers' voices in non-native dialects, especially in pitch-accent languages. CD-TTS is important for developing voice agents that naturally…

text-to-speechText to Speech

Dialect-Specific Models for Automatic Speech Recognition of African American Vernacular English

2019-09-01 · RANLP 2019 9 · Rachel Dorn

African American Vernacular English (AAVE) is a widely-spoken dialect of English, yet it is under-represented in major speech corpora. As a result, speakers of this dialect are often misunderstood by NLP applications. Th…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Voice Adaptation for Swiss German

2025-05-28 · Samuel Stucki, Jan Deriu, Mark Cieliebak

This work investigates the performance of Voice Adaptation models for Swiss German dialects, i.e., translating Standard German text to Swiss German dialect speech. For this, we preprocess a large dataset of Swiss podcast…

Voice Cloning

Sonos Voice Control Bias Assessment Dataset: A Methodology for Demographic Bias Assessment in Voice Assistants

2024-05-14 · Chloé Sekkat, Fanny Leroy, Salima Mdhaffar, Blake Perry Smith 외

Recent works demonstrate that voice assistants do not perform equally well for everyone, but research on demographic robustness of speech technologies is still scarce. This is mainly due to the rarity of large datasets w…

Automatic Speech RecognitionDiversityspeech-recognitionSpeech Recognition+1