paper-with-me

홈 › Papers

Beyond Confidence: Adaptive Abstention in Dual-Threshold Conformal Prediction for Autonomous System Perception

2025-02-11 · Divake Kumar, Nastaran Darabi, Sina Tayebati, Amit Ranjan Trivedi

Safety-critical perception systems require both reliable uncertainty quantification and principled abstention mechanisms to maintain safety under diverse operational conditions. We present a novel dual-threshold conformalization framework that provides statistically-guaranteed uncertainty estimates while enabling selective prediction in high-risk scenarios. Our approach uniquely combines a conformal threshold ensuring valid prediction sets with an abstention threshold optimized through ROC analysis, providing distribution-free coverage guarantees (>= 1 - alpha) while identifying unreliable predictions. Through comprehensive evaluation on CIFAR-100, ImageNet1K, and ModelNet40 datasets, we demonstrate superior robustness across camera and LiDAR modalities under varying environmental perturbations. The framework achieves exceptional detection performance (AUC: 0.993 to 0.995) under severe conditions while maintaining high coverage (>90.0%) and enabling adaptive abstention (13.5% to 63.4% +/- 0.5) as environmental severity increases. For LiDAR-based perception, our approach demonstrates particularly strong performance, maintaining robust coverage (>84.5%) while appropriately abstaining from unreliable predictions. Notably, the framework shows remarkable stability under heavy perturbations, with detection performance (AUC: 0.995 +/- 0.001) significantly outperforming existing methods across all modalities. Our unified approach bridges the gap between theoretical guarantees and practical deployment needs, offering a robust solution for safety-critical autonomous systems operating in challenging real-world conditions.

📄 PDF Abstract BibTeX arXiv:2502.07255

Code (1)

divake/Conformal_Prediction_based_Sensor_Trustworthiness_Detection 공식 구현 pytorch

Tasks

Conformal PredictionUncertainty Quantificationvalid

Similar Papers 제목 키워드 기반

Causal Evidence that Language Models use Confidence to Drive Behavior

2026-03-23 · Dharshan Kumaran, Nathaniel Daw, Simon Osindero, Petar Veličković 외 arxiv

Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstrates that confidence signals can be extracted from language model outputs…

Learning Conformal Abstention Policies for Adaptive Risk Management in Large Language and Vision-Language Models

2025-02-08 · Sina Tayebati, Divake Kumar, Nastaran Darabi, Dinithi Jayasuriya 외

Large Language and Vision-Language Models (LLMs/VLMs) are increasingly used in safety-critical applications, yet their opaque decision-making complicates risk assessment and reliability. Uncertainty quantification (UQ) h…

Conformal PredictionDecision MakingHallucinationInformativeness+4

Explicit Abstention Knobs for Predictable Reliability in Video Question Answering

2025-12-31 · Jorge Ortiz arxiv

High-stakes deployment of vision-language models (VLMs) requires selective prediction, where systems abstain when uncertain rather than risk costly errors. We investigate whether confidence-based abstention provides reli…

Video Question Answering

Eigenmood Space: Uncertainty-Aware Spectral Graph Analysis of Psychological Patterns in Classical Persian Poetry

2026-02-18 · Kourosh Shahnazari, Seyed Moein Ayyoubzadeh, Mohammadali Keshtparvar arxiv

Classical Persian poetry is a historically sustained archive in which affective life is expressed through metaphor, intertextual convention, and rhetorical indirection. These properties make close reading indispensable w…

Improving LLM Reliability through Hybrid Abstention and Adaptive Detection

2026-02-17 · Ankit Sharma, Nachiket Tapas, Jyotiprakash Patra arxiv

Large Language Models (LLMs) deployed in production environments face a fundamental safety-utility trade-off either a strict filtering mechanisms prevent harmful outputs but often block benign queries or a relaxed contro…