paper-with-me

홈 › Papers

Optimistic Feasible Search for Closed-Loop Fair Threshold Decision-Making

2025-12-26 · Wenzhang Du arxiv

Closed-loop decision-making systems (e.g., lending, screening, or recidivism risk assessment) often operate under fairness and service constraints while inducing feedback effects: decisions change who appears in the future, yielding non-stationary data and potentially amplifying disparities. We study online learning of a one-dimensional threshold policy from bandit feedback under demographic parity (DP) and, optionally, service-rate constraints. The learner observes only a scalar score each round and selects a threshold; reward and constraint residuals are revealed only for the chosen threshold. We propose Optimistic Feasible Search (OFS), a simple grid-based method that maintains confidence bounds for reward and constraint residuals for each candidate threshold. At each round, OFS selects a threshold that appears feasible under confidence bounds and, among those, maximizes optimistic reward; if no threshold appears feasible, OFS selects the threshold minimizing optimistic constraint violation. This design directly targets feasible high-utility thresholds and is particularly effective for low-dimensional, interpretable policy classes where discretization is natural. We evaluate OFS on (i) a synthetic closed-loop benchmark with stable contraction dynamics and (ii) two semi-synthetic closed-loop benchmarks grounded in German Credit and COMPAS, constructed by training a score model and feeding group-dependent acceptance decisions back into population composition. Across all environments, OFS achieves higher reward with smaller cumulative constraint violation than unconstrained and primal-dual bandit baselines, and is near-oracle relative to the best feasible fixed threshold under the same sweep procedure. Experiments are reproducible and organized with double-blind-friendly relative outputs.

📄 PDF Abstract BibTeX arXiv:2512.22313

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the Low-Complexity of Fair Learning for Combinatorial Multi-Armed Bandit

2025-01-01 · Xiaoyi Wu, Bo Ji, Bin Li

Combinatorial Multi-Armed Bandit with fairness constraints is a framework where multiple arms form a super arm and can be pulled in each round under uncertainty to maximize cumulative rewards while ensuring the minimum a…

Fairness

KG-FairDiff: Knowledge Graph-Guided Prompt Refinement for Demographically Fair Text-to-Image Generation

2026-05-31 · Farbod Davoodi, Seyed Reza Tavakoli Shiyadeh, Pooria Safaei, Sana Harighi 외 arxiv

Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cultural stereotypes they inherit from training data (rendering women,…

Text-to-Image Generation

Adaptive-Horizon Conflict-Based Search for Closed-Loop Multi-Agent Path Finding

2026-02-12 · Jiarui Li, Federico Pecora, Runyu Zhang, Gioele Zardini arxiv

Multi-Agent Path Finding (MAPF) is a core coordination problem for large robot fleets in automated warehouses and logistics. Existing approaches are typically either open-loop planners, which must compute complete trajec…

Computational Efficiency

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

2026-01-28 · Ruicheng Ao, David Simchi-Levi, Xinshang Wang arxiv

Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), identifying constraint conflicts, and repairing formulations until feasibility …

Hydra-NeXt: Robust Closed-Loop Driving with Open-Loop Training

2025-03-15 · Zhenxin Li, Shihao Wang, Shiyi Lan, Zhiding Yu 외

End-to-end autonomous driving research currently faces a critical challenge in bridging the gap between open-loop training and closed-loop deployment. Current approaches are trained to predict trajectories in an open-loo…

Autonomous DrivingBench2DriveNavSimTrajectory Prediction