paper-with-me

홈 › Papers

How Safe Will I Be Given What I Saw? Calibrated Prediction of Safety Chances for Image-Controlled Autonomy

2025-08-12 · Zhenjiang Mao, Mrinall Eashaan Umasudhan, Ivan Ruchkin arxiv

Autonomous robots that rely on deep neural network controllers pose critical challenges for safety prediction, especially under partial observability and distribution shift. Traditional model-based verification techniques are limited in scalability and require access to low-dimensional state models, while model-free methods often lack reliability guarantees. This paper addresses these limitations by introducing a framework for calibrated safety prediction in end-to-end vision-controlled systems, where neither the state-transition model nor the observation model is accessible. Building on the foundation of world models, we leverage variational autoencoders and recurrent predictors to forecast future latent trajectories from raw image sequences and estimate the probability of satisfying safety properties. We distinguish between monolithic and composite prediction pipelines and introduce a calibration mechanism to quantify prediction confidence. In long-horizon predictions from high-dimensional observations, the forecasted inputs to the safety evaluator can deviate significantly from the training distribution due to compounding prediction errors and changing environmental conditions, leading to miscalibrated risk estimates. To address this, we incorporate unsupervised domain adaptation to ensure robustness of safety evaluation under distribution shift in predictions without requiring manual labels. Our formulation provides theoretical calibration guarantees and supports practical evaluation across long prediction horizons. Experimental results on three benchmarks show that our UDA-equipped evaluators maintain high accuracy and substantially lower false positive rates under distribution shift. Similarly, world model-based composite predictors outperform their monolithic counterparts on long-horizon tasks, and our conformal calibration provides reliable statistical bounds.

📄 PDF Abstract BibTeX arXiv:2508.09346

Code (0)

등록된 구현이 없습니다.

Tasks

Unsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

How Safe Am I Given What I See? Calibrated Prediction of Safety Chances for Image-Controlled Autonomy

2023-08-23 · Zhenjiang Mao, Carson Sobolewski, Ivan Ruchkin

End-to-end learning has emerged as a major paradigm for developing autonomous systems. Unfortunately, with its performance and convenience comes an even greater challenge of safety assurance. A key factor of this challen…

Conformal PredictionPrediction

When to Act: Calibrated Confidence for Reliable Human Intention Prediction in Assistive Robotics

2026-01-08 · Johannes A. Gaus, Winfried Ilg, Daniel Haeufle arxiv

Assistive devices must determine both what a user intends to do and how reliable that prediction is before providing support. We introduce a safety-critical triggering framework based on calibrated probabilities for mult…

When does a predictor know its own loss?

2025-02-27 · Aravind Gollakota, Parikshit Gopalan, Aayush Karan, Charlotte Peale 외

Given a predictor and a loss function, how well can we predict the loss that the predictor will incur on an input? This is the problem of loss prediction, a key computational task associated with uncertainty estimation f…

FairnessPrediction

Medical Model Synthesis Architectures: A Case Study

2026-05-10 · Katherine M. Collins, Marlene Berke, Ilia Sucholutsky, Ayman Ali 외 arxiv

Medicine is rife with high-stakes uncertainty. Doctors routinely make clinical judgments and decisions that juggle many fundamental unknowns, like predictions about what might be causing a patients' symptoms or decisions…

Risk Alignment in Agentic AI Systems

2024-10-02 · Hayley Clatterbuck, Clinton Castro, Arvo Muñoz Morán

Agentic AIs $-$ AIs that are capable and permitted to undertake complex actions with little supervision $-$ mark a new frontier in AI capabilities and raise new questions about how to safely create and align such systems…