paper-with-me

홈 › Papers

Can We Trust AI to Govern AI? Benchmarking LLM Performance on Privacy and AI Governance Exams

2025-08-12 · Zane Witherspoon, Thet Mon Aye, YingYing Hao arxiv

The rapid emergence of large language models (LLMs) has raised urgent questions across the modern workforce about this new technology's strengths, weaknesses, and capabilities. For privacy professionals, the question is whether these AI systems can provide reliable support on regulatory compliance, privacy program management, and AI governance. In this study, we evaluate ten leading open and closed LLMs, including models from OpenAI, Anthropic, Google DeepMind, Meta, and DeepSeek, by benchmarking their performance on industry-standard certification exams: CIPP/US, CIPM, CIPT, and AIGP from the International Association of Privacy Professionals (IAPP). Each model was tested using official sample exams in a closed-book setting and compared to IAPP's passing thresholds. Our findings show that several frontier models such as Gemini 2.5 Pro and OpenAI's GPT-5 consistently achieve scores exceeding the standards for professional human certification - demonstrating substantial expertise in privacy law, technical controls, and AI governance. The results highlight both the strengths and domain-specific gaps of current LLMs and offer practical insights for privacy officers, compliance leads, and technologists assessing the readiness of AI tools for high-stakes data governance roles. This paper provides an overview for professionals navigating the intersection of AI advancement and regulatory risk and establishes a machine benchmark based on human-centric evaluations.

📄 PDF Abstract BibTeX arXiv:2508.09036

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Blockchain-Enabled Approach to Cross-Border Compliance and Trust

2025-01-15 · Vikram Kulothungan

As artificial intelligence (AI) systems become increasingly integral to critical infrastructure and global operations, the need for a unified, trustworthy governance framework is more urgent that ever. This paper propose…

Ethics

From Privacy to Trust in the Agentic Era: A Taxonomy of Challenges in Trustworthy Federated Learning Through the Lens of Trust Report 2.0

2025-07-21 · Nuria Rodríguez-Barroso, Mario García-Márquez, M. Victoria Luzón, Francisco Herrera arxiv

Federated Learning (FL) enables privacy-preserving collaborative learning, yet deployments increasingly show that privacy guarantees alone do not sustain trust in high-risk settings. As FL systems move toward agentic AI,…

Federated Learning

Distributed Machine Learning and the Semblance of Trust

2021-12-21 · Dmitrii Usynin, Alexander Ziller, Daniel Rueckert, Jonathan Passerat-Palmbach 외

The utilisation of large and diverse datasets for machine learning (ML) at scale is required to promote scientific insight into many meaningful problems. However, due to data governance regulations such as GDPR as well a…

BIG-bench Machine LearningFederated LearningPrivacy Preserving

Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps

2026-07-11 · Sakshi Gorkhali, Jonesh Shrestha arxiv

Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally controlled. Federated learning offers a practical alternative by allowing hospita…

Federated Learning

Trustworthy Distributed AI Systems: Robustness, Privacy, and Governance

2024-02-02 · Wenqi Wei, Ling Liu

Emerging Distributed AI systems are revolutionizing big data computing and data processing capabilities with growing economic and societal impact. However, recent studies have identified new attack surfaces and risks cau…

Fairness