paper-with-me

Papers

Automated Phonological Transcription of Akkadian Cuneiform Text

2020-05-01 · LREC 2020 5 · Aleksi Sahala, Miikka Silfverberg, Antti Arppe, Krister Lind{\'e}n

Akkadian was an East-Semitic language spoken in ancient Mesopotamia. The language is attested on hundreds of thousands of cuneiform clay tablets. Several Akkadian text corpora contain only the transliterated text. In this paper, we investigate automated phonological transcription of the transliterated corpora. The phonological transcription provides a linguistically appealing form to represent Akkadian, because the transcription is normalized according to the grammatical description of a given dialect and explicitly shows the Akkadian renderings for Sumerian logograms. Because cuneiform text does not mark the inflection for logograms, the inflected form needs to be inferred from the sentence context. To the best of our knowledge, this is the first documented attempt to automatically transcribe Akkadian. Using a context-aware neural network model, we are able to automatically transcribe syllabic tokens at near human performance with 96{\%} recall @ 3, while the logogram transcription remains more challenging at 82{\%} recall @ 3.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Similar Papers 제목 키워드 기반

Word Segmentation for Akkadian Cuneiform

2016-05-01 · LREC 2016 5 · Timo Homburg, Christian Chiarcos

We present experiments on word segmentation for Akkadian cuneiform, an ancient writing system and a language used for about 3 millennia in the ancient Near East. To our best knowledge, this is the first study of this kin…

Segmentation

Advanced Deep Learning Approaches for Automated Recognition of Cuneiform Symbols

2025-05-07 · Shahad Elshehaby, Alavikunhu Panthakkan, Hussain Al-Ahmad, Mina Al-Saad

This paper presents a thoroughly automated method for identifying and interpreting cuneiform characters via advanced deep-learning algorithms. Five distinct deep-learning models were trained on a comprehensive dataset of…

Deep Learning

Investigating Machine Learning Methods for Language and Dialect Identification of Cuneiform Texts

2020-09-22 · WS 2019 6 · Ehsan Doostmohammadi, Minoo Nassajian

Identification of the languages written using cuneiform symbols is a difficult task due to the lack of resources and the problem of tokenization. The Cuneiform Language Identification task in VarDial 2019 addresses the p…

BIG-bench Machine LearningDialect IdentificationLanguage Identification

TwistBytes - Identification of Cuneiform Languages and German Dialects at VarDial 2019

2019-06-01 · WS 2019 6 · Fern Benites, o, Pius von D{\"a}niken, Mark Cieliebak

We describe our approaches for the German Dialect Identification (GDI) and the Cuneiform Language Identification (CLI) tasks at the VarDial Evaluation Campaign 2019. The goal was to identify dialects of Swiss German in G…

Dialect IdentificationLanguage Identification

Experiments in Cuneiform Language Identification

2019-04-27 · WS 2019 6 · Gustavo Henrique Paetzold, Marcos Zampieri

This paper presents methods to discriminate between languages and dialects written in Cuneiform script, one of the first writing systems in the world. We report the results obtained by the PZ team in the Cuneiform Langua…

Language Identification