Orthographic and Morphological Processing for Persian-to-English Statistical Machine Translation
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationTranslationTransliterationSimilar Papers 제목 키워드 기반
H-Net++: Hierarchical Dynamic Chunking for Tokenizer-Free Language Modelling in Morphologically-Rich Languages
Byte-level language models eliminate fragile tokenizers but face computational challenges in morphologically-rich languages (MRLs), where words span many bytes. We propose H-NET++, a hierarchical dynamic-chunking model t…
Computational EfficiencyLanguage ModellingMIZAN: A Large Persian-English Parallel Corpus
One of the most major and essential tasks in natural language processing is machine translation that is now highly dependent upon multilingual parallel corpora. Through this paper, we introduce the biggest Persian-Englis…
Machine TranslationSentenceTranslationAn Unsupervised Method for Uncovering Morphological Chains
Most state-of-the-art systems today produce morphological analysis based only on orthographic patterns. In contrast, we propose a model for unsupervised morphological analysis that integrates orthographic and semantic vi…
Morphological AnalysisPersian-Spanish Low-Resource Statistical Machine Translation Through English as Pivot Language
This paper is an attempt to exclusively focus on investigating the pivot language technique in which a bridging language is utilized to increase the quality of the Persian-Spanish low-resource Statistical Machine Transla…
Machine TranslationSentenceTranslationSS4MCT: A Statistical Stemmer for Morphologically Complex Texts
There have been multiple attempts to resolve various inflection matching problems in information retrieval. Stemming is a common approach to this end. Among many techniques for stemming, statistical stemming has been sho…
Information RetrievalRetrieval