paper-with-me

Papers

Scalable multilingual PII annotation for responsible AI in LLMs

2025-10-03 · Bharti Meena, Joanna Skubisz, Harshit Rajgarhia, Nand Dave, Kiran Ganesh, Shivali Dalmia, Abhishek Mukherji, Vasudevan Sundarababu arxiv

As Large Language Models (LLMs) gain wider adoption, ensuring their reliable handling of Personally Identifiable Information (PII) across diverse regulatory contexts has become essential. This work introduces a scalable multilingual data curation framework designed for high-quality PII annotation across 13 underrepresented locales, covering approximately 336 locale-specific PII types. Our phased, human-in-the-loop annotation methodology combines linguistic expertise with rigorous quality assurance, leading to substantial improvements in recall and false positive rates from pilot, training, and production phases. By leveraging inter-annotator agreement metrics and root-cause analysis, the framework systematically uncovers and resolves annotation inconsistencies, resulting in high-fidelity datasets suitable for supervised LLM fine-tuning. Beyond reporting empirical gains, we highlight common annotator challenges in multilingual PII labeling and demonstrate how iterative, analytics-driven pipelines can enhance both annotation quality and downstream model reliability.

📄 PDF Abstract BibTeX arXiv:2510.06250

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Soteria: Language-Specific Functional Parameter Steering for Multilingual Safety Alignment

2025-02-16 · Somnath Banerjee, Sayan Layek, Pratyush Chatterjee, Animesh Mukherjee 외

Ensuring consistent safety across multiple languages remains a significant challenge for large language models (LLMs). We introduce Soteria, a lightweight yet powerful strategy that locates and minimally adjusts the "fun…

Safety Alignment

Bias Beyond Borders: Political Ideology Evaluation and Steering in Multilingual LLMs

2026-01-30 · Afrozah Nadeem, Agrima Seth, Mehwish Nasim, Usman Naseem arxiv

Large Language Models (LLMs) increasingly shape global discourse, making fairness and ideological neutrality essential for responsible AI deployment. Despite growing attention to political bias in LLMs, prior work largel…

Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities

2026-05-30 · Wajdi Zaghouani arxiv

Large language models have rapidly evolved in multilingual competence and reasoning capacity, enabling their integration into Social Sciences and Humanities research workflows. Yet existing evaluation paradigms remain an…

Bridging Latent Reasoning and Target-Language Generation via Retrieval-Transition Heads

2026-02-25 · Shaswat Patel, Vishvesh Trivedi, Yue Han, Yihuai Hong 외 arxiv

Recent work has identified a subset of attention heads in Transformer as retrieval heads, which are responsible for retrieving information from the context. In this work, we first investigate retrieval heads in multiling…

XLQA: A Benchmark for Locale-Aware Multilingual Open-Domain Question Answering

2025-08-22 · Keon-Woo Roh, Yeong-Joon Ju, Seong-Whan Lee arxiv

Large Language Models (LLMs) have shown significant progress in Open-domain question answering (ODQA), yet most evaluations focus on English and assume locale-invariant answers across languages. This assumption neglects …

Open-Domain Question Answering