paper-with-me

홈 › Papers

Curriculum-Guided Antifragile Reinforcement Learning for Secure UAV Deconfliction under Observation-Space Attacks

2025-06-26 · Deepak Kumar Panda, Adolfo Perrusquia, Weisi Guo

Reinforcement learning (RL) policies deployed in safety-critical systems, such as unmanned aerial vehicle (UAV) navigation in dynamic airspace, are vulnerable to out-ofdistribution (OOD) adversarial attacks in the observation space. These attacks induce distributional shifts that significantly degrade value estimation, leading to unsafe or suboptimal decision making rendering the existing policy fragile. To address this vulnerability, we propose an antifragile RL framework designed to adapt against curriculum of incremental adversarial perturbations. The framework introduces a simulated attacker which incrementally increases the strength of observation-space perturbations which enables the RL agent to adapt and generalize across a wider range of OOD observations and anticipate previously unseen attacks. We begin with a theoretical characterization of fragility, formally defining catastrophic forgetting as a monotonic divergence in value function distributions with increasing perturbation strength. Building on this, we define antifragility as the boundedness of such value shifts and derive adaptation conditions under which forgetting is stabilized. Our method enforces these bounds through iterative expert-guided critic alignment using Wasserstein distance minimization across incrementally perturbed observations. We empirically evaluate the approach in a UAV deconfliction scenario involving dynamic 3D obstacles. Results show that the antifragile policy consistently outperforms standard and robust RL baselines when subjected to both projected gradient descent (PGD) and GPS spoofing attacks, achieving up to 15% higher cumulative reward and over 30% fewer conflict events. These findings demonstrate the practical and theoretical viability of antifragile reinforcement learning for secure and resilient decision-making in environments with evolving threat scenarios.

📄 PDF Abstract BibTeX arXiv:2506.21129

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

GPS Greedy Policy Search (GPS) is a simple algorithm that learns a policy for test-time data augmentation based on the predictive performance on a validation set. GPS starts with…

Similar Papers 제목 키워드 기반

Robust Policy Switching for Antifragile Reinforcement Learning for UAV Deconfliction in Adversarial Environments

2025-06-26 · Deepak Kumar Panda, Weisi Guo

The increasing automation of navigation for unmanned aerial vehicles (UAVs) has exposed them to adversarial attacks that exploit vulnerabilities in reinforcement learning (RL) through sensor manipulation. Although existi…

Reinforcement Learning (RL)Thompson Sampling

Integrated Conflict Management for UAM with Strategic Demand Capacity Balancing and Learning-based Tactical Deconfliction

2023-05-17 · Shulu Chen, Antony Evans, Marc Brittain, Peng Wei

Urban air mobility (UAM) has the potential to revolutionize our daily transportation, offering rapid and efficient deliveries of passengers and cargo between dedicated locations within and around the urban environment. B…

Managementreinforcement-learningReinforcement Learning

Antifragile Perimeter Control: Anticipating and Gaining from Disruptions with Reinforcement Learning

2024-02-20 · Linghang Sun, Michail A. Makridis, Alexander Genser, Cristian Axenie 외

The optimal operation of transportation systems is often susceptible to unexpected disruptions, such as traffic accidents and social events. Many established control strategies relying on mathematical models can struggle…

Deep Reinforcement LearningModel Predictive ControlReinforcement Learning (RL)

Position: AI Safety Must Embrace an Antifragile Perspective

2025-09-11 · Ming Jin, Hyunin Lee arxiv

This position paper contends that modern AI research must adopt an antifragile perspective on safety -- one in which the system's capacity to guarantee long-term AI safety such as handling rare or out-of-distribution (OO…

SARI: Structured Audio Reasoning via Curriculum-Guided Reinforcement Learning

2025-04-22 · Cheng Wen, Tingwei Guo, Shuaijiang Zhao, Wei Zou 외

Recent work shows that reinforcement learning(RL) can markedly sharpen the reasoning ability of large language models (LLMs) by prompting them to "think before answering." Yet whether and how these gains transfer to audi…

Multiple-choicereinforcement-learningReinforcement LearningReinforcement Learning (RL)