paper-with-me

홈 › Papers

Fatigue-Aware Learning to Defer via Constrained Optimisation

2026-04-01 · Zheng Zhang, Cuong C. Nguyen, David Rosewarne, Kevin Wells, Gustavo Carneiro arxiv

Learning to defer (L2D) enables human-AI cooperation by deciding when an AI system should act autonomously or defer to a human expert. Existing L2D methods, however, assume static human performance, contradicting well-established findings on fatigue-induced degradation. We propose Fatigue-Aware Learning to Defer via Constrained Optimisation (FALCON), which explicitly models workload-varying human performance using psychologically grounded fatigue curves. FALCON formulates L2D as a Constrained Markov Decision Process (CMDP) whose state includes both task features and cumulative human workload, and optimises accuracy under human-AI cooperation budgets via PPO-Lagrangian training. We further introduce FA-L2D, a benchmark that systematically varies fatigue dynamics from near-static to rapidly degrading regimes. Experiments across multiple datasets show that FALCON consistently outperforms state-of-the-art L2D methods across coverage levels, generalises zero-shot to unseen experts with different fatigue patterns, and demonstrates the advantage of adaptive human-AI collaboration over AI-only or human-only decision-making when coverage lies strictly between 0 and 1.

📄 PDF Abstract BibTeX arXiv:2604.00904

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Coverage-Constrained Human-AI Cooperation with Multiple Experts

2024-11-18 · Zheng Zhang, Cuong Nguyen, Kevin Wells, Thanh-Toan Do 외

Human-AI cooperative classification (HAI-CC) approaches aim to develop hybrid intelligent systems that enhance decision-making in various high-stakes real-world scenarios by leveraging both human expertise and AI capabil…

Decision Making

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

2026-06-08 · Emre Turan arxiv

As LLM agents begin to take real, irreversible actions (shell commands, file edits, deploys), the standard safety pattern is a human-in-the-loop approval gate: risky actions pause and wait for a person. We argue the gate…

Coherent Hierarchical Multi-Label Learning to Defer for Medical Imaging

2026-05-04 · Joshua Strong, Pramit Saha, Emma Sun, Helen Higham 외 arxiv

Learning to Defer (L2D) enables a model to predict autonomously or defer to an expert, but prior work largely assumes flat label spaces. We study the first L2D setting with hierarchical multi-label decisions, motivated b…

Multi-Label Learning

MPD$^2$-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis

2026-05-08 · Wenxin Zhan arxiv

Learning-to-defer (L2D) can make glaucoma screening safer by routing difficult/uncertain cases to humans, yet standard formulations overlook expert availability, heterogeneous readers behavior, workload imbalance, asymme…

A Hybrid Hinge-Beam Continuum Robot with Passive Safety Capping for Real-Time Fatigue Awareness

2025-09-11 · Tongshun Chen, Zezhou Sun, Yanhan Sun, Yuhao Wang 외 arxiv

Cable-driven continuum robots offer high flexibility and lightweight design, making them well-suited for tasks in constrained and unstructured environments. However, prolonged use can induce mechanical fatigue from plast…