Towards Producing Human-Validated Translation Resources for the Fula language through WordNet Linking
We propose methods to link automatically parsed linguistic data to the WordNet. We apply these methods on a trilingual dictionary in Fula, English and French. Dictionary entry parsing is used to collect the linguistic data. Then we connect it to the Open Multilingual WordNet (OMW) through two attempts, and use confidence scores to quantify accuracy. We obtained 11,000 entries in parsing and linked about 58{\%} to the OMW on the first attempt, and an additional 14{\%} in the second one. These links are due to be validated by Fula speakers before being added to the Kamusi Project{'}s database.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationTranslationSimilar Papers 제목 키워드 기반
HERDPhobia: A Dataset for Hate Speech against Fulani in Nigeria
Social media platforms allow users to freely share their opinions about issues or anything they feel like. However, they also make it easier to spread hate and abusive content. The Fulani ethnic group has been the victim…
MindfulAgents: Personalizing Mindfulness Meditation via an Expert-Aligned Multi-Agent System
Mindfulness meditation is a widely accessible and evidence-based method for supporting mental health. Despite the proliferation of mindfulness meditation apps, sustaining user engagement remains a persistent challenge. P…
Model Stitching by Functional Latent Alignment
Evaluating functional similarity involves quantifying the degree to which independently trained neural networks learn functionally similar representations. Reliably inferring the functional similarity of these networks r…
Knowledge DistillationmodelBootstrapping Niche Multilingual Code Translation via Reinforcement Learning with Execution-Based Verifiable Supervision
Code translation must preserve executable behavior across many programming languages, yet neural code translation has largely focused on a few popular languages such as C++, Java, and Python. This leaves a niche, many-to…
Reinforcement LearningCode TranslationLarge-Scale, Diverse, Paraphrastic Bitexts via Sampling and Clustering
Producing diverse paraphrases of a sentence is a challenging task. Natural paraphrase corpora are scarce and limited, while existing large-scale resources are automatically generated via back-translation and rely on beam…
ClusteringDiversitySentenceTranslation