paper-with-me

Papers

A Validity Perspective on Evaluating the Justified Use of Data-driven Decision-making Algorithms

2022-06-30 · Amanda Coston, Anna Kawakami, Haiyi Zhu, Ken Holstein, Hoda Heidari

Recent research increasingly brings to question the appropriateness of using predictive tools in complex, real-world tasks. While a growing body of work has explored ways to improve value alignment in these tools, comparatively less work has centered concerns around the fundamental justifiability of using these tools. This work seeks to center validity considerations in deliberations around whether and how to build data-driven algorithms in high-stakes domains. Toward this end, we translate key concepts from validity theory to predictive algorithms. We apply the lens of validity to re-examine common challenges in problem formulation and data issues that jeopardize the justifiability of using predictive algorithms and connect these challenges to the social science discourse around validity. Our interdisciplinary exposition clarifies how these concepts apply to algorithmic decision making contexts. We demonstrate how these validity considerations could distill into a series of high-level questions intended to promote and document reflections on the legitimacy of the predictive task and the suitability of the data.

📄 PDF Abstract BibTeX arXiv:2206.14983

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Declare and Justify: Explicit assumptions in AI evaluations are necessary for effective regulation

2024-11-19 · Peter Barnett, Lisa Thiergart

As AI systems advance, AI evaluations are becoming an important pillar of regulations for ensuring safety. We argue that such regulation should require developers to explicitly identify and justify key underlying assumpt…

Litmus: Zero-Label, Code-Driven Metric Specification for Evaluating AI Systems

2026-06-22 · Prajjwal Gupta, Prasang Gupta, Vishal Bhutani, Apoorva Sharma 외 arxiv

As agentic LLM systems move from prototypes to deployment across increasingly diverse domains, evaluating them has become both more important and more difficult. The challenge is not only that individual metrics may be u…

Architectural Sweet Spots for Modeling Human Label Variation by the Example of Argument Quality: It's Best to Relate Perspectives!

2023-11-06 · Philipp Heinisch, Matthias Orlikowski, Julia Romberg, Philipp Cimiano

Many annotation tasks in natural language processing are highly subjective in that there can be different valid and justified perspectives on what is a proper label for a given example. This also applies to the judgment …

Recommendation Systemsvalid

Evaluating LLM-Driven Summarisation of Parliamentary Debates with Computational Argumentation

2026-04-21 · Eoghan Cunningham, Derek Greene, James Cross, Antonio Rago arxiv

Understanding how policy is debated and justified in parliament is a fundamental aspect of the democratic process. However, the volume and complexity of such debates mean that outside audiences struggle to engage. Meanwh…

Overview of the 2022 Validity and Novelty Prediction Shared Task

2022-10-01 · ArgMining (ACL) 2022 10 · Philipp Heinisch, Anette Frank, Juri Opitz, Moritz Plenz 외

This paper provides an overview of the Argument Validity and Novelty Prediction Shared Task that was organized as part of the 9th Workshop on Argument Mining (ArgMining 2022). The task focused on the prediction of the va…

Argument MiningBinary ClassificationPredictionValNov