paper-with-me

Papers

Normalized Orthography for Tunisian Arabic

2024-02-20 · Houcemeddine Turki, Kawthar Ellouze, Hager Ben Ammar, Mohamed Ali Hadj Taieb, Imed Adel, Mohamed Ben Aouicha, Pier Luigi Farri, Abderrezak Bennour

Tunisian Arabic (ISO 693-3: aeb) isa distinct variety native to Tunisia, derived from Arabic and enriched by various historical influences. This research introduces the "Normalized Orthography for Tunisian Arabic" (NOTA), an adaptation of CODA* guidelines for transcribing Tunisian Arabic using Arabic script. The aim is to enhance language resource development by ensuring user-friendliness and consistency. The updated standard addresses challenges in accurately representing Tunisian phonology and morphology, correcting issues from transcriptions based on Modern Standard Arabic.

📄 PDF Abstract BibTeX arXiv:2402.12940

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Conventional Orthography for Tunisian Arabic

2014-05-01 · LREC 2014 5 · In{\`e}s Zribi, Rahma Boujelbane, Abir Masmoudi, Mariem ellouze 외

Tunisian Arabic is a dialect of the Arabic language spoken in Tunisia. Tunisian Arabic is an under-resourced language. It has neither a standard orthography nor large collections of written text and dictionaries. Actuall…

Language ModellingMachine TranslationSpeech RecognitionSpeech Synthesis+1

Multi-Task Sequence Prediction For Tunisian Arabizi Multi-Level Annotation

2020-11-10 · COLING (WANLP) 2020 12 · Elisa Gugliotta, Marco Dinarelli, Olivier Kraif

In this paper we propose a multi-task sequence prediction system, based on recurrent neural networks and used to annotate on multiple levels an Arabizi Tunisian corpus. The annotation performed are text classification, t…

POSPOS Taggingtext-classificationText Classification

LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect

2025-04-03 · Hedi Naouara, Jean-Pierre Lorré, Jérôme Louradour

Developing Automatic Speech Recognition (ASR) systems for Tunisian Arabic Dialect is challenging due to the dialect's linguistic complexity and the scarcity of annotated speech datasets. To address these challenges, we p…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

Sentiment Analysis of Tunisian Dialects: Linguistic Ressources and Experiments

2017-04-01 · WS 2017 4 · Salima Medhaffar, Fethi Bougares, Yannick Est{\`e}ve, Lamia Hadrich-Belguith

Dialectal Arabic (DA) is significantly different from the Arabic language taught in schools and used in written communication and formal speech (broadcast news, religion, politics, etc.). There are many existing research…

Sentiment Analysis

Parallel resources for Tunisian Arabic Dialect Translation

2020-12-01 · COLING (WANLP) 2020 12 · Saméh Kchaou, Rahma Boujelbane, Lamia Hadrich-Belguith

The difficulty of processing dialects is clearly observed in the high cost of building representative corpus, in particular for machine translation. Indeed, all machine translation systems require a huge amount and good …

Data AugmentationMachine TranslationManagementSentence+1