Natural Language Generation for Polysynthetic Languages: Language Teaching and Learning Software for Kanyen'k\'eha (Mohawk)
Kanyen{'}k{\'e}ha (in English, Mohawk) is an Iroquoian language spoken primarily in Eastern Canada (Ontario, Qu{\'e}bec). Classified as endangered, it has only a small number of speakers and very few younger native speakers. Consequently, teachers and courses, teaching materials and software are urgently needed. In the case of software, the polysynthetic nature of Kanyen{'}k{\'e}ha means that the number of possible combinations grows exponentially and soon surpasses attempts to capture variant forms by hand. It is in this context that we describe an attempt to produce language teaching materials based on a generative approach. A natural language generation environment (ivi/Vinci) embedded in a web environment (VinciLingua) makes it possible to produce, by rule, variant forms of indefinite complexity. These may be used as models to explore, or as materials to which learners respond. Generated materials may take the form of written text, oral utterances, or images; responses may be typed on a keyboard, gestural (using a mouse) or, to a limited extent, oral. The software also provides complex orthographic, morphological and syntactic analysis of learner productions. We describe the trajectory of development of materials for a suite of four courses on Kanyen{'}k{\'e}ha, the first of which will be taught in the fall of 2018.
Code (0)
등록된 구현이 없습니다.
Tasks
Text GenerationSimilar Papers 제목 키워드 기반
Unsupervised Morphological Segmentation for Low-Resource Polysynthetic Languages
Polysynthetic languages pose a challenge for morphological analysis due to the root-morpheme complexity and to the word class {``}squish{''}. In addition, many of these polysynthetic languages are low-resource. We propos…
Morphological AnalysisLost in Translation: Analysis of Information Loss During Machine Translation Between Polysynthetic and Fusional Languages
Machine translation from polysynthetic to fusional languages is a challenging task, which gets further complicated by the limited amount of parallel text available. Thus, translation performance is far from the state of …
Machine TranslationTranslationNeural Polysynthetic Language Modelling
Research in natural language processing commonly assumes that approaches that work well for English and and other widely-used languages are "language agnostic". In high-resource languages, especially those that are analy…
Language ModellingLemmatizationMachine TranslationComputational Challenges for Polysynthetic Languages
Given advances in computational linguistic analysis of complex languages using Machine Learning as well as standard Finite State Transducers, coupled with recent efforts in language revitalization, the time was right to …
BIG-bench Machine LearningExpanding Universal Dependencies for Polysynthetic Languages: A Case of St. Lawrence Island Yupik
This paper describes the development of the first Universal Dependencies (UD) treebank for St. Lawrence Island Yupik, an endangered language spoken in the Bering Strait region. While the UD guidelines provided a general …
Dependency Parsing