paper-with-me

홈 › Papers

Knowledge Beyond Language: Bridging the Gap in Multilingual Machine Unlearning Evaluation

2026-05-14 · Kyomin Hwang, Hyeonjin Kim, Sangyeon Cho, Nojun Kwak arxiv

While LLMs are increasingly used in commercial services, they pose privacy risks such as leakage of sensitive personally identifiable information (PII). For LLMs trained on multilingual corpora, Multilingual Machine Unlearning (MMU) aims to remove information across multiple languages. However, prior MMU evaluations fail to capture such cross-linguistic distribution of information, being largely limited to direct extensions of per-language evaluation protocols. To this end, we propose two metrics to evaluate the information spread across languages: the Knowledge Separability Score (KSS) and the Knowledge Persistence Score (KPS). KSS measures the overall unlearning quality across multiple languages, while KPS more specifically aims to assess consistent removal of information among different language pairs. We evaluated various unlearning methods in the multilingual setting with these metrics and conducted comprehensive analyses. Through our investigation, we provide insights into unique phenomena exclusive to MMU and offer a new perspective on MMU evaluation.

📄 PDF Abstract BibTeX arXiv:2605.14404

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Cross-lingual Word Embeddings beyond Zero-shot Machine Translation

2020-11-03 · Shifei Chen, Ali Basirat

We explore the transferability of a multilingual neural machine translation model to unseen languages when the transfer is grounded solely on the cross-lingual word embeddings. Our experimental results show that the tran…

Cross-Lingual Word EmbeddingsMachine TranslationTranslationWord Embeddings+1

Increasing Coverage and Precision of Textual Information in Multilingual Knowledge Graphs

2023-11-27 · Simone Conia, Min Li, Daniel Lee, Umar Farooq Minhas 외

Recent work in Natural Language Processing and Computer Vision has been using textual information -- e.g., entity names and descriptions -- available in knowledge graphs to ground neural models to high-quality structured…

Entity LinkingKnowledge Graph CompletionKnowledge GraphsMachine Translation+1

Marco-LLM: Bridging Languages via Massive Multilingual Training for Cross-Lingual Enhancement

2024-12-05 · Lingfeng Ming, Bo Zeng, Chenyang Lyu, Tianqi Shi 외

Large Language Models (LLMs) have achieved remarkable progress in recent years; however, their excellent performance is still largely limited to major world languages, primarily English. Many LLMs continue to face challe…

BelebeleMachine Translation

The Multilingual Divide and Its Impact on Global AI Safety

2025-05-27 · Aidan Peppin, Julia Kreutzer, Alice Schoenauer Sebag, Kelly Marchisio 외

Despite advances in large language model capabilities in recent years, a large gap remains in their capabilities and safety performance for many languages beyond a relatively small handful of globally dominant languages.…

Language ModelingLanguage ModellingLarge Language Model

CONCAP: Seeing Beyond English with Concepts Retrieval-Augmented Captioning

2025-07-27 · George Ibrahim, Rita Ramos, Yova Kementchedjhieva arxiv

Multilingual vision-language models have made significant strides in image captioning, yet they still lag behind their English counterparts due to limited multilingual training data and costly large-scale model parameter…

Image Captioning