paper-with-me

Papers

Investigating Machine Learning Methods for Language and Dialect Identification of Cuneiform Texts

2020-09-22 · WS 2019 6 · Ehsan Doostmohammadi, Minoo Nassajian

Identification of the languages written using cuneiform symbols is a difficult task due to the lack of resources and the problem of tokenization. The Cuneiform Language Identification task in VarDial 2019 addresses the problem of identifying seven languages and dialects written in cuneiform; Sumerian and six dialects of Akkadian language: Old Babylonian, Middle Babylonian Peripheral, Standard Babylonian, Neo-Babylonian, Late Babylonian, and Neo-Assyrian. This paper describes the approaches taken by SharifCL team to this problem in VarDial 2019. The best result belongs to an ensemble of Support Vector Machines and a naive Bayes classifier, both working on character-level features, with macro-averaged F1-score of 72.10%.

📄 PDF Abstract BibTeX arXiv:2009.10794

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningDialect IdentificationLanguage Identification

Similar Papers 제목 키워드 기반

Towards spoken dialect identification of Irish

2023-07-14 · Liam Lonergan, Mengjie Qian, Neasa Ní Chiaráin, Christer Gobl 외

The Irish language is rich in its diversity of dialects and accents. This compounds the difficulty of creating a speech recognition system for the low-resource language, as such a system must contend with a high degree o…

Dialect IdentificationLanguage Identificationspeech-recognitionSpeech Recognition

Automatic Arabic Dialect Identification Systems for Written Texts: A Survey

2020-09-26 · Maha J. Althobaiti

Arabic dialect identification is a specific task of natural language processing, aiming to automatically predict the Arabic dialect of a given text. Arabic dialect identification is the first step in various natural lang…

Dialect IdentificationMachine TranslationSentenceSpeech Synthesis+6

Demographic Dialectal Variation in Social Media: A Case Study of African-American English

2016-08-31 · EMNLP 2016 11 · Su Lin Blodgett, Lisa Green, Brendan O'Connor

Though dialectal language is increasingly abundant on social media, few resources exist for developing NLP tools to handle such language. We conduct a case study of dialectal language in online conversational text by inv…

Dependency ParsingLanguage Identification

Comparing Pipelined and Integrated Approaches to Dialectal Arabic Neural Machine Translation

2019-06-01 · WS 2019 6 · Pamela Shapiro, Kevin Duh

When translating diglossic languages such as Arabic, situations may arise where we would like to translate a text but do not know which dialect it is. A traditional approach to this problem is to design dialect identific…

Dialect IdentificationMachine TranslationTranslation

The Curious Case of Logistic Regression for Italian Languages and Dialects Identification

2022-10-01 · VarDial (COLING) 2022 10 · Giacomo Camposampiero, Quynh Anh Nguyen, Francesco Di Stefano

Automatic Language Identification represents an important task for improving many real-world applications such as opinion mining and machine translation. In the case of closely-related languages such as regional dialects…

Language IdentificationMachine TranslationOpinion Miningregression+1