paper-with-me

Papers

System-level Safety Guard: Safe Tracking Control through Uncertain Neural Network Dynamics Models

2023-12-11 · Xiao Li, Yutong Li, Anouck Girard, Ilya Kolmanovsky

The Neural Network (NN), as a black-box function approximator, has been considered in many control and robotics applications. However, difficulties in verifying the overall system safety in the presence of uncertainties hinder the deployment of NN modules in safety-critical systems. In this paper, we leverage the NNs as predictive models for trajectory tracking of unknown dynamical systems. We consider controller design in the presence of both intrinsic uncertainty and uncertainties from other system modules. In this setting, we formulate the constrained trajectory tracking problem and show that it can be solved using Mixed-integer Linear Programming (MILP). The proposed MILP-based approach is empirically demonstrated in robot navigation and obstacle avoidance through simulations. The demonstration videos are available at https://xiaolisean.github.io/publication/2023-11-01-L4DC2024.

📄 PDF Abstract BibTeX arXiv:2312.06810

Code (1)

xiaolisean/milpsafetyguard 공식 구현

Tasks

Robot Navigation

Similar Papers 제목 키워드 기반

RSafe: Incentivizing proactive reasoning to build robust and adaptive LLM safeguards

2025-06-09 · Jingnan Zheng, Xiangtian Ji, Yijun Lu, Chenhang Cui 외

Large Language Models (LLMs) continue to exhibit vulnerabilities despite deliberate safety alignment efforts, posing significant risks to users and society. To safeguard against the risk of policy-violating content, syst…

Safety Alignment

AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models

2026-04-10 · Mintong Kang, Chen Fang, Bo Li arxiv

Audio has rapidly become a primary interface for foundation models, powering real-time voice assistants. Ensuring safety in audio systems is inherently more complex than just "unsafe text spoken aloud": real-world risks …

Red Teaming

SalamahBench: Toward Standardized Safety Evaluation for Arabic Language Models

2026-02-03 · Omar Abdelnasser, Fatemah Alharbi, Khaled Khasawneh, Ihsen Alouani 외 arxiv

Safety alignment in Language Models (LMs) is fundamental for trustworthy AI. However, while different stakeholders are trying to leverage Arabic Language Models (ALMs), systematic safety evaluation of ALMs remains largel…

Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning

2024-10-31 · Nabil Omi, Hosein Hasanbeig, Hiteshi Sharma, Sriram K. Rajamani 외

In this paper we propose a formal, model-agnostic meta-learning framework for safe reinforcement learning. Our framework is inspired by how parents safeguard their children across a progression of increasingly riskier ta…

Meta-LearningMinecraftreinforcement-learningReinforcement Learning+1

Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content

2026-05-28 · Ihor Stepanov, Aleksandr Smechov arxiv

Real-time safety filtering for large language model (LLM) applications requires classifiers that can detect unsafe prompts, toxic language, jailbreak attempts, and unsafe responses without the cost profile of large guard…