paper-with-me

Papers

What Makes an Evaluation Useful? Common Pitfalls and Best Practices

2025-03-30 · Gil Gekker, Meirav Segal, Dan Lahav, Omer Nevo

Following the rapid increase in Artificial Intelligence (AI) capabilities in recent years, the AI community has voiced concerns regarding possible safety risks. To support decision-making on the safe use and development of AI systems, there is a growing need for high-quality evaluations of dangerous model capabilities. While several attempts to provide such evaluations have been made, a clear definition of what constitutes a "good evaluation" has yet to be agreed upon. In this practitioners' perspective paper, we present a set of best practices for safety evaluations, drawing on prior work in model evaluation and illustrated through cybersecurity examples. We first discuss the steps of the initial thought process, which connects threat modeling to evaluation design. Then, we provide the characteristics and parameters that make an evaluation useful. Finally, we address additional considerations as we move from building specific evaluations to building a full and comprehensive evaluation suite.

📄 PDF Abstract BibTeX arXiv:2503.23424

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Decoding FL Defenses: Systemization, Pitfalls, and Remedies

2025-02-03 · Momin Ahmad Khan, Virat Shejwalkar, Yasra Chandio, Amir Houmansadr 외

While the community has designed various defenses to counter the threat of poisoning attacks in Federated Learning (FL), there are no guidelines for evaluating these defenses. These defenses are prone to subtle pitfalls …

Federated LearningSurvey

AI for the Common Good?! Pitfalls, challenges, and Ethics Pen-Testing

2018-10-30 · Bettina Berendt

Recently, many AI researchers and practitioners have embarked on research visions that involve doing AI for "Good". This is part of a general drive towards infusing AI research and practice with ethical thinking. One fre…

Ethics

The Devil is in the Details: On the Pitfalls of Event Extraction Evaluation

2023-06-12 · Hao Peng, Xiaozhi Wang, Feng Yao, Kaisheng Zeng 외

Event extraction (EE) is a crucial task aiming at extracting events from texts, which includes two subtasks: event detection (ED) and event argument extraction (EAE). In this paper, we check the reliability of EE evaluat…

Event Argument ExtractionEvent DetectionEvent Extraction

Five Pitfalls When Assessing Synthetic Medical Images with Reference Metrics

2024-08-12 · Melanie Dohmen, Tuan Truong, Ivo M. Baltruschat, Matthias Lenga

Reference metrics have been developed to objectively and quantitatively compare two images. Especially for evaluating the quality of reconstructed or compressed images, these metrics have shown very useful. Extensive tes…

SSIM

Pitfalls of Assessing Extracted Hierarchies for Multi-Class Classification

2021-01-26 · Pablo del Moral, Slawomir Nowaczyk, Anita Sant'Anna, Sepideh Pashami

Using hierarchies of classes is one of the standard methods to solve multi-class classification problems. In the literature, selecting the right hierarchy is considered to play a key role in improving classification perf…

ClassificationGeneral ClassificationMulti-class Classification