paper-with-me

홈 › Papers

Position: AI Safety Must Embrace an Antifragile Perspective

2025-09-11 · Ming Jin, Hyunin Lee arxiv

This position paper contends that modern AI research must adopt an antifragile perspective on safety -- one in which the system's capacity to guarantee long-term AI safety such as handling rare or out-of-distribution (OOD) events expands over time. Conventional static benchmarks and single-shot robustness tests overlook the reality that environments evolve and that models, if left unchallenged, can drift into maladaptation (e.g., reward hacking, over-optimization, or atrophy of broader capabilities). We argue that an antifragile approach -- Rather than striving to rapidly reduce current uncertainties, the emphasis is on leveraging those uncertainties to better prepare for potentially greater, more unpredictable uncertainties in the future -- is pivotal for the long-term reliability of open-ended ML systems. In this position paper, we first identify key limitations of static testing, including scenario diversity, reward hacking, and over-alignment. We then explore the potential of antifragile solutions to manage rare events. Crucially, we advocate for a fundamental recalibration of the methods used to measure, benchmark, and continually improve AI safety over the long term, complementing existing robustness approaches by providing ethical and practical guidelines towards fostering an antifragile AI safety community.

📄 PDF Abstract BibTeX arXiv:2509.13339

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Antifragile Control Systems: The case of an oscillator-based network model of urban road traffic dynamics

2022-10-19 · Cristian Axenie, Margherita Grossi

Existing traffic control systems only possess a local perspective over the multiple scales of traffic evolution, namely the intersection level, the corridor level, and the region level respectively. But luckily, despite …

Curriculum-Guided Antifragile Reinforcement Learning for Secure UAV Deconfliction under Observation-Space Attacks

2025-06-26 · Deepak Kumar Panda, Adolfo Perrusquia, Weisi Guo

Reinforcement learning (RL) policies deployed in safety-critical systems, such as unmanned aerial vehicle (UAV) navigation in dynamic airspace, are vulnerable to out-ofdistribution (OOD) adversarial attacks in the observ…

Decision MakingReinforcement Learning (RL)

Antifragile control systems in neuronal processing: A sensorimotor perspective

2024-04-23 · Cristian Axenie

The stability--robustness--resilience--adaptiveness continuum in neuronal processing follows a hierarchical structure that explains interactions and information processing among the different time scales. Interestingly, …

Dynamic Rate Splitting Grouping for Antifragile Responses to Wireless Network Disruptions

2024-05-13 · Kevin Weinberger, Aydin Sezgin

The reliance on wireless network architectures for applications demanding high reliability and fault tolerance is growing. These architectures heavily depend on wireless channels, making them susceptible to impairments a…

Antifragile Perimeter Control: Anticipating and Gaining from Disruptions with Reinforcement Learning

2024-02-20 · Linghang Sun, Michail A. Makridis, Alexander Genser, Cristian Axenie 외

The optimal operation of transportation systems is often susceptible to unexpected disruptions, such as traffic accidents and social events. Many established control strategies relying on mathematical models can struggle…

Deep Reinforcement LearningModel Predictive ControlReinforcement Learning (RL)