paper-with-me

Papers

Towards Safe Multilingual Frontier AI

2024-09-06 · Artūrs Kanepajs, Vladimir Ivanov, Richard Moulange

Linguistically inclusive LLMs -- which maintain good performance regardless of the language with which they are prompted -- are necessary for the diffusion of AI benefits around the world. Multilingual jailbreaks that rely on language translation to evade safety measures undermine the safe and inclusive deployment of AI systems. We provide policy recommendations to enhance the multilingual capabilities of AI while mitigating the risks of multilingual jailbreaks. We examine how a language's level of resourcing relates to how vulnerable LLMs are to multilingual jailbreaks in that language. We do this by testing five advanced AI models across 24 official languages of the EU. Building on prior research, we propose policy actions that align with the EU legal landscape and institutional framework to address multilingual jailbreaks, while promoting linguistic inclusivity. These include mandatory assessments of multilingual capabilities and vulnerabilities, public opinion research, and state support for multilingual AI development. The measures aim to improve AI safety and functionality through EU policy initiatives, guiding the implementation of the EU AI Act and informing regulatory efforts of the European AI Office.

📄 PDF Abstract BibTeX arXiv:2409.13708

Code (1)

akanepajs/multilingual 공식 구현

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

K-EXAONE 2.0 Technical Report

2026-08-05 · Eunbi Choi, Kibong Choi, Sehyun Chun, Seokhee Hong 외 hf

This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward global frontier-scale foundation models. Rather than training from scra…

Long-Context Understanding

A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5

2026-01-15 · Xingjun Ma, Yixu Wang, Hengyuan Xu, Yutao Wu 외 arxiv

The rapid evolution of Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs) has driven major gains in reasoning, perception, and generation across language and vision, yet whether these advances tran…

Adversarial RobustnessImage Generation

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss

2026-04-14 · Ronald Skorobogat, Ameya Prabhu, Matthias Bethge arxiv

Multilingual benchmarks guide the development of frontier models. Yet multilingual evaluations reported by frontier models are structured similar to popular reasoning and knowledge benchmarks, but across many languages. …

Mathematical Reasoning

XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity

2026-05-07 · Dasol Choi, Eugenia Kim, Jaewon Noh, Sang Seo 외 arxiv

Current LLM safety benchmarks are predominantly English-centric and often rely on translation, failing to capture country-specific harms. Moreover, they rarely evaluate a model's ability to detect culturally embedded sen…

Improving Methodologies for LLM Evaluations Across Global Languages

2026-01-22 · Akriti Vij, Benjamin Chua, Darshini Ramiah, En Qi Ng 외 arxiv

As frontier AI models are deployed globally, it is essential that their behaviour remains safe and reliable across diverse linguistic and cultural contexts. To examine how current model safeguards hold up in such setting…