paper-with-me

홈 › Papers

Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages

2026-02-14 · Somnath Banerjee, Rima Hazra, Animesh Mukherjee arxiv

Large language models (LLMs) are being deployed across the Global South, where everyday use involves low-resource languages, code-mixing, and culturally specific norms. Yet safety pipelines, benchmarks, and alignment still largely target English and a handful of high-resource languages, implicitly assuming safety and factuality ''transfer'' across languages. Evidence increasingly shows they do not. We synthesize recent findings indicating that (i) safety guardrails weaken sharply on low-resource and code-mixed inputs, (ii) culturally harmful behavior can persist even when standard toxicity scores look acceptable, and (iii) English-only knowledge edits and safety patches often fail to carry over to low-resource languages. In response, we outline a practical agenda for researchers and students in the Global South: parameter-efficient safety steering, culturally grounded evaluation and preference data, and participatory workflows that empower local communities to define and mitigate harm. Our aim is to make multilingual safety a core requirement-not an add-on-for equitable AI in underrepresented regions.

📄 PDF Abstract BibTeX arXiv:2602.13867

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CultureGuard: Towards Culturally-Aware Dataset and Guard Model for Multilingual Safety Applications

2025-08-03 · Raviraj Joshi, Rakesh Paul, Kanishk Singla, Anusha Kamath 외 arxiv

The increasing use of Large Language Models (LLMs) in agentic applications highlights the need for robust safety guard models. While content safety in English is well-studied, non-English languages lack similar advanceme…

Synthetic Data GenerationZero-shot GeneralizationCross-Lingual TransferMachine Translation

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

2026-03-18 · Priyaranjan Pattnayak, Sanchari Chowdhuri arxiv

As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages remains poorly understood. We present the first systematic evaluation of LLM safe…

SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia

2026-02-02 · Panuthep Tasawong, Jian Gang Ngui, Alham Fikri Aji, Trevor Cohn 외 arxiv

Culturally aware safeguards are crucial for AI alignment in real-world settings, where safety extends beyond common sense and encompasses diverse local values, norms, and region-specific regulations. However, building la…

Machine Translation

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models

2026-05-01 · Yunhan Zhao, Zhaorun Chen, Xingjun Ma, Yu-Gang Jiang 외 arxiv

As Large Language Models (LLMs) are increasingly deployed in cross-linguistic contexts, ensuring safety in diverse regulatory and cultural environments has become a critical challenge. However, existing multilingual benc…

Machine Translation

The Multilingual Divide and Its Impact on Global AI Safety

2025-05-27 · Aidan Peppin, Julia Kreutzer, Alice Schoenauer Sebag, Kelly Marchisio 외

Despite advances in large language model capabilities in recent years, a large gap remains in their capabilities and safety performance for many languages beyond a relatively small handful of globally dominant languages.…

Language ModelingLanguage ModellingLarge Language Model