Toward an NLG System for Bantu languages: first steps with Runyankore (demo)
There are many domain-specific and language-specific NLG systems, of which it may be possible to adapt to related domains and languages. The languages in the Bantu language family have their own set of features distinct from other major groups, which therefore severely limits the options to bootstrap an NLG system from existing ones. We present here our first proof-of-concept application for knowledge-to-text NLG as a plugin to the Protege 5.x ontology development system, tailored to Runyankore, a Bantu language indigenous to Uganda. It comprises a basic annotation model for linguistic information such as noun class, an implementation of existing verbalisation rules and a CFG for verbs, and a basic interface for data entry.
Code (0)
등록된 구현이 없습니다.
Tasks
Text GenerationSimilar Papers 제목 키워드 기반
Noun Class Disambiguation in Runyankore and Related Languages
Bantu languages are spoken by communities in more than half of the countries on the African continent by an estimated third of a billion people. Despite this populous and the amount of high quality linguistic research do…
Pluralizing Nouns across Agglutinating Bantu Languages
Text generation may require the pluralization of nouns, such as in context-sensitive user interfaces and in natural language generation more broadly. While this has been solved for the widely-used languages, this is stil…
Text GenerationEvaluation of a Runyankore grammar engine for healthcare messages
Natural Language Generation (NLG) can be used to generate personalized health information, which is especially useful when provided in one{'}s own language. However, the NLG technique widely used in different domains and…
Machine TranslationText GenerationTowards Computational Resource Grammars for Runyankore and Rukiga
In this paper, we present computational resource grammars of Runyankore and Rukiga (R{\&}R) languages. Runyankore and Rukiga are two under-resourced Bantu Languages spoken by about 6 million people indigenous to South- W…
DescriptiveGenerating Varied Training Corpora in Runyankore Using a Combined Semantic and Syntactic, Pattern-Grammar-based Approach
Machine learning algorithms have been applied to achieve high levels of accuracy in tasks associated with the processing of natural language. However, these algorithms require large amounts of training data in order to p…
BIG-bench Machine LearningSentiment AnalysisWord EmbeddingsWord Similarity