The Adaptive Stress Testing Formulation
Validation is a key challenge in the search for safe autonomy. Simulations are often either too simple to provide robust validation, or too complex to tractably compute. Therefore, approximate validation methods are needed to tractably find failures without unsafe simplifications. This paper presents the theory behind one such black-box approach: adaptive stress testing (AST). We also provide three examples of validation problems formulated to work with AST.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Adaptive Stress Testing of Trajectory Predictions in Flight Management Systems
To find failure events and their likelihoods in flight-critical systems, we investigate the use of an advanced black-box stress testing approach called adaptive stress testing. We analyze a trajectory predictor from a de…
Decision MakingManagementSequential Decision MakingAdaptive Stress Testing with Reward Augmentation for Autonomous Vehicle Validation
Determining possible failure scenarios is a critical step in the evaluation of autonomous vehicle systems. Real-world vehicle testing is commonly employed for autonomous vehicle validation, but the costs and time require…
Reinforcement LearningAdaptive Stress Testing: Finding Likely Failure Events with Reinforcement Learning
Finding the most likely path to a set of failure states is important to the analysis of safety-critical systems that operate over a sequence of time steps, such as aircraft collision avoidance systems and autonomous cars…
Autonomous DrivingCollision Avoidancereinforcement-learningReinforcement Learning+1Adaptive Stress Testing for Adversarial Learning in a Financial Environment
We demonstrate the use of Adaptive Stress Testing to detect and address potential vulnerabilities in a financial environment. We develop a simplified model for credit card fraud detection that utilizes a linear regressio…
Fraud Detectionregressionreinforcement-learningReinforcement Learning (RL)Adaptive Stress Testing Black-Box LLM Planners
Large language models (LLMs) have recently demonstrated success in generalizing across decision-making tasks including planning, control and prediction, but their tendency to hallucinate unsafe and undesired outputs pose…