paper-with-me

Papers

Towards Inclusive NLP: Assessing Compressed Multilingual Transformers across Diverse Language Benchmarks

2025-07-25 · Maitha Alshehhi, Ahmed Sharshar, Mohsen Guizani arxiv

Although LLMs have attained significant success in high-resource languages, their capacity in low-resource linguistic environments like Kannada and Arabic is not yet fully understood. This work benchmarking the performance of multilingual and monolingual Large Language Models (LLMs) across Arabic, English, and Indic languages, with particular emphasis on the effects of model compression strategies such as pruning and quantization. Findings shows significant performance differences driven by linguistic diversity and resource availability on SOTA LLMS as BLOOMZ, AceGPT, Jais, LLaMA-2, XGLM, and AraGPT2. We find that multilingual versions of the model outperform their language-specific counterparts across the board, indicating substantial cross-lingual transfer benefits. Quantization (4-bit and 8-bit) is effective in maintaining model accuracy while promoting efficiency, but aggressive pruning significantly compromises performance, especially in bigger models. Our findings pinpoint key strategies to construct scalable and fair multilingual NLP solutions and underscore the need for interventions to address hallucination and generalization errors in the low-resource setting.

📄 PDF Abstract BibTeX arXiv:2507.19699

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual TransferModel Compression

Similar Papers 제목 키워드 기반

Probing Gender Bias in Multilingual LLMs: A Case Study of Stereotypes in Persian

2025-09-24 · Ghazal Kalhor, Behnam Bahrak arxiv

Multilingual Large Language Models (LLMs) are increasingly used worldwide, making it essential to ensure they are free from gender bias to prevent representational harm. While prior studies have examined such biases in h…

Towards Safe Multilingual Frontier AI

2024-09-06 · Artūrs Kanepajs, Vladimir Ivanov, Richard Moulange

Linguistically inclusive LLMs -- which maintain good performance regardless of the language with which they are prompted -- are necessary for the diffusion of AI benefits around the world. Multilingual jailbreaks that re…

Inclusive Artificial Intelligence

2022-12-24 · Dilip Arumugam, Shi Dong, Benjamin Van Roy

Prevailing methods for assessing and comparing generative AIs incentivize responses that serve a hypothetical representative individual. Evaluating models in these terms presumes homogeneous preferences across the popula…

The ML-SUPERB 2.0 Challenge: Towards Inclusive ASR Benchmarking for All Language Varieties

2025-09-08 · William Chen, Chutong Meng, Jiatong Shi, Martijn Bartelds 외 arxiv

Recent improvements in multilingual ASR have not been equally distributed across languages and language varieties. To advance state-of-the-art (SOTA) ASR models, we present the Interspeech 2025 ML-SUPERB 2.0 Challenge. W…

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

2025-11-14 · Jian Gao, Richeng Xuan, Zhaolu Kang, Dingshi Liao 외 arxiv

The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian languages like Lao. To fill this gap, we introduce \textbf{LaoBench}, t…