paper-with-me

홈 › Papers

Mind the Uncertainty: Risk-Aware and Actively Exploring Model-Based Reinforcement Learning

2023-09-11 · Marin Vlastelica, Sebastian Blaes, Cristina Pineri, Georg Martius

We introduce a simple but effective method for managing risk in model-based reinforcement learning with trajectory sampling that involves probabilistic safety constraints and balancing of optimism in the face of epistemic uncertainty and pessimism in the face of aleatoric uncertainty of an ensemble of stochastic neural networks.Various experiments indicate that the separation of uncertainties is essential to performing well with data-driven MPC approaches in uncertain and safety-critical control environments.

📄 PDF Abstract BibTeX arXiv:2309.05582

Code (0)

등록된 구현이 없습니다.

Tasks

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

2026-04-10 · Zukun Zhang, Kai Shu, Mingqiao Mo arxiv

Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise, and unstructured contact conditions. We introduce SafeMind, a differe…

On Surprising Effects of Risk-Aware Domain Randomization for Contact-Rich Sampling-based Predictive Control

2026-05-05 · Sergio A. Esteban, Junheng Li, Vince Kurtz, Aaron D. Ames arxiv

Domain randomization (DR) is widely used in policy learning to improve robustness to modeling error, but remains underexplored in contact-rich sampling-based predictive control (SPC), where rollout quality is highly sens…

The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models

2025-06-15 · Peiyuan Tang, Haojie Xin, Xiaodong Zhang, Jun Sun 외

As Vision-Language Models (VLMs) demonstrate increasing capabilities across real-world applications such as code generation and chatbot assistance, ensuring their safety has become paramount. Unlike traditional Large Lan…

ChatbotCode GenerationText Generation

EthicMind: A Risk-Aware Framework for Ethical-Emotional Alignment in Multi-Turn Dialogue

2026-04-10 · Jiawen Deng, Wei Li, Wentao Zhang, Ziyun Jiao 외 arxiv

Intelligent dialogue systems are increasingly deployed in emotionally and ethically sensitive settings, where failures in either emotional attunement or ethical judgment can cause significant harm. Existing dialogue mode…

TableMind++: An Uncertainty-Aware Programmatic Agent for Tool-Augmented Table Reasoning

2026-03-08 · Mingyue Cheng, Shuo Yu, Chuang Jiang, Xiaoyu Tao 외 arxiv

Table reasoning requires models to jointly perform semantic understanding and precise numerical operations. Most existing methods rely on a single-turn reasoning paradigm over tables which suffers from context overflow a…

Reinforcement Learning