paper-with-me

홈 › Papers

Sociotechnical Safety Evaluation of Generative AI Systems

2023-10-18 · Laura Weidinger, Maribeth Rauh, Nahema Marchal, Arianna Manzini, Lisa Anne Hendricks, Juan Mateos-Garcia, Stevie Bergman, Jackie Kay, Conor Griffin, Ben Bariach, Iason Gabriel, Verena Rieser, William Isaac

Generative AI systems produce a range of risks. To ensure the safety of generative AI systems, these risks must be evaluated. In this paper, we make two main contributions toward establishing such evaluations. First, we propose a three-layered framework that takes a structured, sociotechnical approach to evaluating these risks. This framework encompasses capability evaluations, which are the main current approach to safety evaluation. It then reaches further by building on system safety principles, particularly the insight that context determines whether a given capability may cause harm. To account for relevant context, our framework adds human interaction and systemic impacts as additional layers of evaluation. Second, we survey the current state of safety evaluation of generative AI systems and create a repository of existing evaluations. Three salient evaluation gaps emerge from this analysis. We propose ways forward to closing these gaps, outlining practical steps as well as roles and responsibilities for different actors. Sociotechnical safety evaluation is a tractable approach to the robust and comprehensive safety evaluation of generative AI systems.

📄 PDF Abstract BibTeX arXiv:2310.11986

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AI red-teaming is a sociotechnical challenge: on values, labor, and harms

2024-12-12 · Tarleton Gillespie, Ryland Shaw, Mary L. Gray, Jina Suh

As generative AI technologies find more and more real-world applications, the importance of testing their performance and safety seems paramount. "Red-teaming" has quickly become the primary approach to test AI models--p…

Red Teaming

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

2026-04-22 · Rebecca L. Johnson arxiv

In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: benchmarks do not just measure, they shape what models appear to be. Fu…

Concrete Safety for ML Problems: System Safety for ML Development and Assessment

2023-02-06 · Edgar W. Jatho, Logan O. Mailloux, Eugene D. Williams, Patrick McClure 외

Many stakeholders struggle to make reliances on ML-driven systems due to the risk of harm these systems may cause. Concerns of trustworthiness, unintended social harms, and unacceptable social and ethical violations unde…

System Safety Engineering for Social and Ethical ML Risks: A Case Study

2022-11-08 · Edgar W. Jatho III, Logan O. Mailloux, Shalaleh Rismani, Eugene D. Williams 외

Governments, industry, and academia have undertaken efforts to identify and mitigate harms in ML-driven systems, with a particular focus on social and ethical risks of ML components in complex sociotechnical systems. How…

Unsafe at any AUC: Unlearned Lessons from Sociotechnical Disasters for Responsible AI

2026-07-15 · Joshua A. Kroll, Andrew Smart, R. Stuart Geiger, Abigail Z. Jacobs arxiv

As automated decision-making and data-driven technologies pervade society and are used to manage consequential outcomes, understanding the technology's capabilities, limitations, and attendant risks in context requires a…