paper-with-me

홈 › Papers

Extracting Incentives from Black-Box Decisions

2019-10-13 · Yonadav Shavit, William S. Moses

An algorithmic decision-maker incentivizes people to act in certain ways to receive better decisions. These incentives can dramatically influence subjects' behaviors and lives, and it is important that both decision-makers and decision-recipients have clarity on which actions are incentivized by the chosen model. While for linear functions, the changes a subject is incentivized to make may be clear, we prove that for many non-linear functions (e.g. neural networks, random forests), classical methods for interpreting the behavior of models (e.g. input gradients) provide poor advice to individuals on which actions they should take. In this work, we propose a mathematical framework for understanding algorithmic incentives as the challenge of solving a Markov Decision Process, where the state includes the set of input features, and the reward is a function of the model's output. We can then leverage the many toolkits for solving MDPs (e.g. tree-based planning, reinforcement learning) to identify the optimal actions each individual is incentivized to take to improve their decision under a given model. We demonstrate the utility of our method by estimating the maximally-incentivized actions in two real-world settings: a recidivism risk predictor we train using ProPublica's COMPAS dataset, and an online credit scoring tool published by the Fair Isaac Corporation (FICO).

📄 PDF Abstract BibTeX arXiv:1910.05664

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Can Desegregation Close the Racial Gap in High School Coursework?

2022-08-25 · Ritika Sethi

This paper examines the interplay between desegregation, institutional bias, and individual behavior in education. Using a game-theoretic model that considers race-heterogeneous social incentives, the study investigates …

Vocal Bursts Intensity Prediction

Measuring an artificial intelligence agent's trust in humans using machine incentives

2022-12-27 · Tim Johnson, Nick Obradovich

Scientists and philosophers have debated whether humans can trust advanced artificial intelligence (AI) agents to respect humanity's best interests. Yet what about the reverse? Will advanced AI agents trust humans? Gaugi…

AI AgentLanguage ModellingLarge Language Model

Interpreting Blackbox Models via Model Extraction

2017-05-23 · Osbert Bastani, Carolyn Kim, Hamsa Bastani

Interpretability has become incredibly important as machine learning is increasingly used to inform consequential decisions. We propose to construct global explanations of complex, blackbox models in the form of a decisi…

modelModel extraction

Segregation Dynamics with Reinforcement Learning and Agent Based Modeling

2019-09-18 · Egemen Sert, Yaneer Bar-Yam, Alfredo J. Morales

Societies are complex. Properties of social systems can be explained by the interplay and weaving of individual actions. Incentives are key to understand people's choices and decisions. For instance, individual preferenc…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents

2026-03-26 · Fitsum Sileshi Beyene, Christopher L. Dancy arxiv

Optical character recognition (OCR) and document understanding systems increasingly rely on large vision and vision-language models, yet evaluation remains centered on modern, Western, and institutional documents. This e…