paper-with-me

홈 › Papers

Implicit Discourse Relation Classification For Nigerian Pidgin

2024-06-26 · Muhammed Saeed, Peter Bourgonje, Vera Demberg

Despite attempts to make Large Language Models multi-lingual, many of the world's languages are still severely under-resourced. This widens the performance gap between NLP and AI applications aimed at well-financed, and those aimed at less-resourced languages. In this paper, we focus on Nigerian Pidgin (NP), which is spoken by nearly 100 million people, but has comparatively very few NLP resources and corpora. We address the task of Implicit Discourse Relation Classification (IDRC) and systematically compare an approach translating NP data to English and then using a well-resourced IDRC tool and back-projecting the labels versus creating a synthetic discourse corpus for NP, in which we translate PDTB and project PDTB labels, and then train an NP IDR classifier. The latter approach of learning a "native" NP classifier outperforms our baseline by 13.27\% and 33.98\% in f$_{1}$ score for 4-way and 11-way classification, respectively.

📄 PDF Abstract BibTeX arXiv:2406.18776

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationImplicit Discourse Relation ClassificationRelationRelation Classification

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Semi-automatic discourse annotation in a low-resource language: Developing a connective lexicon for Nigerian Pidgin

2021-11-01 · CODI 2021 11 · Marian Marchal, Merel Scholman, Vera Demberg

Cross-linguistic research on discourse structure and coherence marking requires discourse-annotated corpora and connective lexicons in a large number of languages. However, the availability of such resources is limited, …

Relation

Semantic Enrichment of Nigerian Pidgin English for Contextual Sentiment Classification

2020-03-27 · Wuraola Fisayo Oyewusi, Olubayo Adekanmbi, Olalekan Akinsande

Nigerian English adaptation, Pidgin, has evolved over the years through multi-language code switching, code mixing and linguistic adaptation. While Pidgin preserves many of the words in the normal English language corpus…

ClassificationGeneral ClassificationSentiment AnalysisSentiment Classification

The Register Gap: A Meaning Intelligence Framework for Nigerian Public Discourse

2026-06-18 · Celestine Achi arxiv

We introduce the Meaning Intelligence Framework (MIF), a nine-dimension annotation and evaluation schema for Nigerian public discourse that separates surface sentiment from true communicative intent. Existing benchmarks …

Towards Supervised and Unsupervised Neural Machine Translation Baselines for Nigerian Pidgin

2020-03-27 · Orevaoghene Ahia, Kelechi Ogueji

Nigerian Pidgin is arguably the most widely spoken language in Nigeria. Variants of this language are also spoken across West and Central Africa, making it a very important language. This work aims to establish supervise…

Machine TranslationNMTTranslation

Low-Resource Cross-Lingual Adaptive Training for Nigerian Pidgin

2023-07-01 · Pin-Jie Lin, Muhammed Saeed, Ernie Chang, Merel Scholman

Developing effective spoken language processing systems for low-resource languages poses several challenges due to the lack of parallel data and limited resources for fine-tuning models. In this work, we target on improv…

text-classificationText ClassificationTranslation