paper-with-me

Papers

Falsifying Discriminant Validity of Predictive Algorithms

2026-01-23 · Amanda Coston arxiv

Empirical investigations into unintended model behavior often show that the algorithm is predicting another outcome than what was intended. These exposés highlight the need to identify when algorithms predict unintended quantities - ideally before deploying them into consequential settings. We propose a falsification framework that provides a principled statistical test for discriminant validity: the requirement that an algorithm predict intended outcomes better than impermissible ones. Drawing on falsification practices from causal inference, econometrics, and psychometrics, our framework compares calibrated prediction losses across outcomes to assess whether the algorithm exhibits discriminant validity with respect to a specified impermissible proxy. In settings where the target outcome is difficult to observe, multiple permissible proxy outcomes may be available; our framework accommodates both this setting and the case with a single permissible proxy. Throughout we use nonparametric hypothesis testing methods that make minimal assumptions on the data-generating process. We illustrate the method in an admissions setting, where the framework establishes discriminant validity with respect to gender but fails to establish discriminant validity with respect to race. This demonstrates how falsification can serve as an early validity check. We also provide analysis in a criminal justice setting, where we highlight the limitations of our framework and emphasize the need for complementary approaches to assess other aspects of construct validity and external validity.

📄 PDF Abstract BibTeX arXiv:2601.17146

Code (0)

등록된 구현이 없습니다.

Tasks

Causal Inference

Similar Papers 제목 키워드 기반

A Validity Perspective on Evaluating the Justified Use of Data-driven Decision-making Algorithms

2022-06-30 · Amanda Coston, Anna Kawakami, Haiyi Zhu, Ken Holstein 외

Recent research increasingly brings to question the appropriateness of using predictive tools in complex, real-world tasks. While a growing body of work has explored ways to improve value alignment in these tools, compar…

Decision Making

AI Psychometrics: Evaluating the Psychological Reasoning of Large Language Models with Psychometric Validities

2026-03-11 · Yibai Li, Xiaolin Lin, Zhenghui Sha, Zhiye Jin 외 arxiv

The immense number of parameters and deep neural networks make large language models (LLMs) rival the complexity of human brains, which also makes them opaque ``black box'' systems that are challenging to evaluate and in…

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

2026-05-08 · Baishi Li, Ta Yu, Kelvin J. L. Koa, Ke-Wei Huang arxiv

Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddings to measure latent constructs such as novelty, creativity, and bia…

Representation Learning

Fast semi-supervised discriminant analysis for binary classification of large data-sets

2017-09-14 · Joris Tavernier, Jaak Simm, Karl Meerbergen, Joerg Kurt Wegner 외

High-dimensional data requires scalable algorithms. We propose and analyze three scalable and related algorithms for semi-supervised discriminant analysis (SDA). These methods are based on Krylov subspace methods which e…

Binary ClassificationGeneral Classificationsubspace methods

Data Strategies for Fleetwide Predictive Maintenance

2018-12-11 · David Noever

For predictive maintenance, we examine one of the largest public datasets for machine failures derived along with their corresponding precursors as error rates, historical part replacements, and sensor inputs. To simplif…

Feature Importanceregression