paper-with-me

Papers

From Script to Semantics: Prompting Strategies for African NLI

2026-06-02 · Anuj Tiwari, Terry Oko-odion, Hannah Nwokocha arxiv

Large language models (LLMs) are increasingly evaluated in multilingual settings, yet their inference behavior in low-resource African languages remains underexplored especially under pure prompting without fine-tuning. We present a systematic study of prompting strategies for Natural Language Inference (NLI) in Swahili, Yoruba, and Hausa using the AfriXNLI benchmark. We evaluate five prompting strategies Baseline (zero-shot), Script-Aware, Language Specific, Contrastive, and Native-Label Self-Translation (NL-STP) across two mid-sized open weight models (Llama3.2-3B and Gemma3-4B). To isolate the effect of prompt design, the effect of few-shot examples and Chain-of-Thought reasoning is eliminated in our study. We find a significant difference in performance of class wise across strategies with highly neutral class collapse and high prediction skew in some configurations. Contrastive prompting proves to be the most reliable and steadily improving strategy over language and model and has better balance of class behavior and balance of overall accuracy gains. Notably, well-constructed prompts are sufficient to beat more powerful baselines that are provided with few-shot prompts and Chain-of-Thought prompts. We have found that prompt formulation is essential to multilingual NLI with low-resource languages and that language aware decision structuring can be used to meaningfully enhance robustness in resource challenged settings.

📄 PDF Abstract BibTeX arXiv:2606.03304

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Language Inference

Similar Papers 제목 키워드 기반

TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages

2026-05-31 · Victor Akinode, Senyu Li, Wassim Hamidouche, Waqas Zamir 외 arxiv

Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, critically underexplored. We introduce TUKABENCH, a jailbreak benchmark…

Towards conversational assistants for health applications: using ChatGPT to generate conversations about heart failure

2025-05-06 · Anuja Tayal, Devika Salunke, Barbara Di Eugenio, Paula G Allen-Meares 외

We explore the potential of ChatGPT (3.5-turbo and 4) to generate conversations focused on self-care strategies for African-American heart failure patients -- a domain with limited specialized datasets. To simulate patie…

Semantics of body parts in African WordNet: a case of Northern Sotho

2016-01-01 · GWC 2016 1 · Mampaka Lydia Mojapelo

This paper presents a linguistic account of the lexical semantics of body parts in African WordNet, with special reference to Northern Sotho. It focuses on external human body parts synsets in Northern Sotho. The paper s…

CapEnrich: Enriching Caption Semantics for Web Images via Cross-modal Pre-trained Knowledge

2022-11-17 · Linli Yao, Weijing Chen, Qin Jin

Automatically generating textual descriptions for massive unlabeled images on the web can greatly benefit realistic web applications, e.g. multimodal retrieval and recommendation. However, existing models suffer from the…

Concept AlignmentRetrieval

Prompt Engineering for Responsible Generative AI Use in African Education: A Report from a Three-Day Training Series

2026-01-04 · Benjamin Quarshie, Vanessa Willemse, Macharious Nabang, Bismark Nyaaba Akanzire 외 arxiv

Generative artificial intelligence (GenAI) tools are increasingly adopted in education, yet many educators lack structured guidance on responsible and context sensitive prompt engineering, particularly in African and oth…

Prompt Engineering