paper-with-me

Papers

To Err Is AI! Debugging as an Intervention to Facilitate Appropriate Reliance on AI Systems

2024-09-22 · Gaole He, Abri Bharos, Ujwal Gadiraju

Powerful predictive AI systems have demonstrated great potential in augmenting human decision making. Recent empirical work has argued that the vision for optimal human-AI collaboration requires 'appropriate reliance' of humans on AI systems. However, accurately estimating the trustworthiness of AI advice at the instance level is quite challenging, especially in the absence of performance feedback pertaining to the AI system. In practice, the performance disparity of machine learning models on out-of-distribution data makes the dataset-specific performance feedback unreliable in human-AI collaboration. Inspired by existing literature on critical thinking and a critical mindset, we propose the use of debugging an AI system as an intervention to foster appropriate reliance. In this paper, we explore whether a critical evaluation of AI performance within a debugging setting can better calibrate users' assessment of an AI system and lead to more appropriate reliance. Through a quantitative empirical study (N = 234), we found that our proposed debugging intervention does not work as expected in facilitating appropriate reliance. Instead, we observe a decrease in reliance on the AI system after the intervention -- potentially resulting from an early exposure to the AI system's weakness. We explore the dynamics of user confidence and user estimation of AI trustworthiness across groups with different performance levels to help explain how inappropriate reliance patterns occur. Our findings have important implications for designing effective interventions to facilitate appropriate reliance and better human-AI collaboration.

📄 PDF Abstract BibTeX arXiv:2409.14377

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Knowing About Knowing: An Illusion of Human Competence Can Hinder Appropriate Reliance on AI Systems

2023-01-25 · Gaole He, Lucie Kuiper, Ujwal Gadiraju

The dazzling promises of AI systems to augment humans in various tasks hinge on whether humans can appropriately rely on them. Recent research has shown that appropriate reliance is the key to achieving complementary tea…

Decision Making

Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance

2025-02-18 · Tejas Srinivasan, Jesse Thomason

Trust biases how users rely on AI recommendations in AI-assisted decision-making tasks, with low and high levels of trust resulting in increased under- and over-reliance, respectively. We propose that AI assistants shoul…

Decision Making

The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs

2025-06-23 · Muntasir Adnan, Carlos C. N. Kuhn

The effectiveness of AI debugging follows a predictable exponential decay pattern; most models lose 60-80% of their debugging capability within just 2-3 attempts, despite iterative debugging being a critical capability f…

Code Generation

DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems

2025-12-07 · Ming Ma, Jue Zhang, Fangkai Yang, Yu Kang 외 arxiv

Large language model (LLM)-based multi-agent systems are challenging to debug because failures often arise from long, branching interaction traces. The prevailing practice is to leverage LLMs for log-based failure locali…

Auto Debugging

Fine-Grained Appropriate Reliance: Human-AI Collaboration with a Multi-Step Transparent Decision Workflow for Complex Task Decomposition

2025-01-19 · Gaole He, Patrick Hemmer, Michael Vössing, Max Schemmer 외

In recent years, the rapid development of AI systems has brought about the benefits of intelligent services but also concerns about security and reliability. By fostering appropriate user reliance on an AI system, both c…

Decision MakingFact CheckingFact Verification