paper-with-me

Papers

Dallah: A Dialect-Aware Multimodal Large Language Model for Arabic

2024-07-25 · Fakhraddin Alwajih, Gagan Bhatia, Muhammad Abdul-Mageed

Recent advancements have significantly enhanced the capabilities of Multimodal Large Language Models (MLLMs) in generating and understanding image-to-text content. Despite these successes, progress is predominantly limited to English due to the scarcity of high quality multimodal resources in other languages. This limitation impedes the development of competitive models in languages such as Arabic. To alleviate this situation, we introduce an efficient Arabic multimodal assistant, dubbed Dallah, that utilizes an advanced language model based on LLaMA-2 to facilitate multimodal interactions. Dallah demonstrates state-of-the-art performance in Arabic MLLMs. Through fine-tuning six Arabic dialects, Dallah showcases its capability to handle complex dialectal interactions incorporating both textual and visual elements. The model excels in two benchmark tests: one evaluating its performance on Modern Standard Arabic (MSA) and another specifically designed to assess dialectal responses. Beyond its robust performance in multimodal interaction tasks, Dallah has the potential to pave the way for further development of dialect-aware Arabic MLLMs.

📄 PDF Abstract BibTeX arXiv:2407.18129

Code (0)

등록된 구현이 없습니다.

Tasks

Image to textLanguage ModelingLanguage ModellingLarge Language Modelmultimodal interactionMultimodal Large Language Model

Similar Papers 제목 키워드 기반

DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation

2025-10-16 · Yu Zhou, Sohyun An, Haikang Deng, Da Yin 외 arxiv

Contact languages like English exhibit rich regional variations in the form of dialects, which are often used by dialect speakers interacting with generative models. However, can multimodal generative models effectively …

multimodal generation

5-Dialects-BN: Unmasking the Impact of Transliteration on Bangla Dialectal LLMs

2026-09-09 · Md Mahir Jawad, Galib Mahmud Jim, Rafid Ahmed, Mir Sazzat Hossain 외 arxiv

Large Language Models (LLMs) have achieved remarkable progress across natural language processing (NLP) tasks, yet their capabilities degrade sharply for low-resource languages and dialectally diverse settings. Bangla, t…

parameter-efficient fine-tuningMachine Translation

Maastricht University at AMIYA: Adapting LLMs for Dialectal Arabic using Fine-tuning and MBR Decoding

2026-02-10 · Abdulhai Alali, Abderrahmane Issam arxiv

Large Language Models (LLMs) are becoming increasingly multilingual, supporting hundreds of languages, especially high resource ones. Unfortunately, Dialect variations are still underrepresented due to limited data and l…

Towards Lexical Encoding of Multi-Word Expressions in Spanish Dialects

2016-05-01 · LREC 2016 5 · Diana Bogantes, Eric Rodr{\'\i}guez, Alej Arauco, Alej Rodr{\'\i}guez 외

This paper describes a pilot study in lexical encoding of multi-word expressions (MWEs) in 4 Latin American dialects of Spanish: Costa Rican, Colombian, Mexican and Peruvian. We describe the variability of MWE usage acro…

A Novel Dialect-Aware Framework for the Classification of Arabic Dialects and Emotions

2025-02-13 · Nasser A Alsadhan

Arabic is one of the oldest languages still in use today. As a result, several Arabic-speaking regions have developed dialects that are unique to them. Dialect and emotion recognition have various uses in Arabic text ana…

Emotion ClassificationEmotion Recognition