Benchmark for Out-of-Distribution Detection in Deep Reinforcement Learning
Reinforcement Learning (RL) based solutions are being adopted in a variety of domains including robotics, health care and industrial automation. Most focus is given to when these solutions work well, but they fail when presented with out of distribution inputs. RL policies share the same faults as most machine learning models. Out of distribution detection for RL is generally not well covered in the literature, and there is a lack of benchmarks for this task. In this work we propose a benchmark to evaluate OOD detection methods in a Reinforcement Learning setting, by modifying the physical parameters of non-visual standard environments or corrupting the state observation for visual environments. We discuss ways to generate custom RL environments that can produce OOD data, and evaluate three uncertainty methods for the OOD detection task. Our results show that ensemble methods have the best OOD detection performance with a lower standard deviation across multiple environments.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement LearningOut-of-Distribution DetectionOut of Distribution (OOD) Detectionreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Out-of-Distribution Dynamics Detection: RL-Relevant Benchmarks and Results
We study the problem of out-of-distribution dynamics (OODD) detection, which involves detecting when the dynamics of a temporal process change compared to the training-distribution dynamics. This is relevant to applicati…
Reinforcement Learning (RL)Time SeriesTime Series AnalysisRethinking Out-of-Distribution Detection for Reinforcement Learning: Advancing Methods for Evaluation and Detection
While reinforcement learning (RL) algorithms have been successfully applied across numerous sequential decision-making problems, their generalization to unforeseen testing environments remains a significant concern. In t…
Out-of-Distribution DetectionOut of Distribution (OOD) DetectionReinforcement Learning (RL)Sequential Decision Making+1OOD-RL-Bench: A Benchmark Framework for Out-of-Distribution Detection in Reinforcement Learning
Reliable reinforcement learning (RL) agents must maintain operational integrity amidst sensor malfunctions, dynamic disturbances, and slow environmental shifts. The detection of out-of-distribution conditions is pivotal …
Out-of-Distribution DetectionReinforcement LearningSCOPED: Score-Curvature Out-of-distribution Proximity Evaluator for Diffusion
Out-of-distribution (OOD) detection is essential for reliable deployment of machine learning systems in vision, robotics, reinforcement learning, and beyond. We introduce Score-Curvature Out-of-distribution Proximity Eva…
Reinforcement LearningDensity EstimationOutlier DetectionOmni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection
Multimodal deepfakes are proliferating on social media and threaten authenticity, information integrity, and digital forensics. Existing benchmarks are constrained by their single-modality scope, simplified manipulations…
DeepFake Detection