paper-with-me

홈 › Papers

The Multilingual Divide and Its Impact on Global AI Safety

2025-05-27 · Aidan Peppin, Julia Kreutzer, Alice Schoenauer Sebag, Kelly Marchisio, Beyza Ermis, John Dang, Samuel Cahyawijaya, Shivalika Singh, Seraphina Goldfarb-Tarrant, Viraat Aryabumi, Aakanksha, Wei-Yin Ko, Ahmet Üstün, Matthias Gallé, Marzieh Fadaee, Sara Hooker

Despite advances in large language model capabilities in recent years, a large gap remains in their capabilities and safety performance for many languages beyond a relatively small handful of globally dominant languages. This paper provides researchers, policymakers and governance experts with an overview of key challenges to bridging the "language gap" in AI and minimizing safety risks across languages. We provide an analysis of why the language gap in AI exists and grows, and how it creates disparities in global AI safety. We identify barriers to address these challenges, and recommend how those working in policy and governance can help address safety concerns associated with the language gap by supporting multilingual dataset creation, transparency, and research.

📄 PDF Abstract BibTeX arXiv:2505.21344

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages

2026-02-14 · Somnath Banerjee, Rima Hazra, Animesh Mukherjee arxiv

Large language models (LLMs) are being deployed across the Global South, where everyday use involves low-resource languages, code-mixing, and culturally specific norms. Yet safety pipelines, benchmarks, and alignment sti…

One Anchor for All: Unified Multilingual and Multimodal Safety Alignment for LVLMs

2026-07-30 · Enyi Shi, Fei Shen, Chuancheng Shi, Linxia Zhu 외 arxiv

As large vision-language models (LVLMs) are deployed globally, the combination of multilingual instructions and visual information makes malicious attacks more covert and sophisticated than ever before. However, existing…

All Languages Matter: On the Multilingual Safety of Large Language Models

2023-10-02 · Wenxuan Wang, Zhaopeng Tu, Chang Chen, Youliang Yuan 외

Safety lies at the core of developing and deploying large language models (LLMs). However, previous safety benchmarks only concern the safety in one language, e.g. the majority language in the pretraining data such as En…

AllSafety Alignment

The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning

2026-06-27 · Will Hawkins, Kaivalya Rawal, Jonathan Rystrøm, Stratis Tsirtsis 외 arxiv

Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown that this increase in capability comes with a cost: it can increase a mo…

LinguaSafe: A Comprehensive Multilingual Safety Benchmark for Large Language Models

2025-08-18 · Zhiyuan Ning, Tianle Gu, Jiaxin Song, Shixin Hong 외 arxiv

The widespread adoption and increasing prominence of large language models (LLMs) in global technologies necessitate a rigorous focus on ensuring their safety across a diverse range of linguistic and cultural contexts. T…