Combining Phonology and Morphology for the Normalization of Historical Texts
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationOptical Character Recognition (OCR)Similar Papers 제목 키워드 기반
Historical German Text Normalization Using Type- and Token-Based Language Modeling
2024-09-04
· Anton Ehrmanntraut
Historic variations of spelling poses a challenge for full-text search or natural language processing on historical digitized texts. To minimize the gap between the historic orthography and contemporary spelling, usually…
DecoderLanguage ModelingLanguage ModellingLarge Language Model+2Normalized Orthography for Tunisian Arabic
2024-02-20
· Houcemeddine Turki, Kawthar Ellouze, Hager Ben Ammar, Mohamed Ali Hadj Taieb 외
Tunisian Arabic (ISO 693-3: aeb) isa distinct variety native to Tunisia, derived from Arabic and enriched by various historical influences. This research introduces the "Normalized Orthography for Tunisian Arabic" (NOTA)…
Proceedings of the 16th Workshop on Computational Research in Phonetics, Phonology, and Morphology
2019-08-01 · WS 2019 8
·
Proceedings of the Fifteenth Workshop on Computational Research in Phonetics, Phonology, and Morphology
2018-10-01 · WS 2018 10
·