paper-with-me

홈 › Papers

Free Energy Risk Metrics for Systemically Safe AI: Gatekeeping Multi-Agent Study

2025-02-06 · Michael Walters, Rafael Kaufmann, Justice Sefas, Thomas Kopinski

We investigate the Free Energy Principle as a foundation for measuring risk in agentic and multi-agent systems. From these principles we introduce a Cumulative Risk Exposure metric that is flexible to differing contexts and needs. We contrast this to other popular theories for safe AI that hinge on massive amounts of data or describing arbitrarily complex world models. In our framework, stakeholders need only specify their preferences over system outcomes, providing straightforward and transparent decision rules for risk governance and mitigation. This framework naturally accounts for uncertainty in both world model and preference model, allowing for decision-making that is epistemically and axiologically humble, parsimonious, and future-proof. We demonstrate this novel approach in a simplified autonomous vehicle environment with multi-agent vehicles whose driving policies are mediated by gatekeepers that evaluate, in an online fashion, the risk to the collective safety in their neighborhood, and intervene through each vehicle's policy when appropriate. We show that the introduction of gatekeepers in an AV fleet, even at low penetration, can generate significant positive externalities in terms of increased system safety.

📄 PDF Abstract BibTeX arXiv:2502.04249

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Identifying systemically important companies in the entire liability network of a small open economy

2018-01-31

To a large extent, the systemic importance of financial institutions is related to the topology of financial liability networks. In this work we reconstruct and analyze the - to our knowledge - largest financial network …

Towards the Quantification of Safety Risks in Deep Neural Networks

2020-09-13 · Peipei Xu, Wenjie Ruan, Xiaowei Huang

Safety concerns on the deep neural networks (DNNs) have been raised when they are applied to critical sectors. In this paper, we define safety risks by requesting the alignment of the network's decision with human percep…

Evaluation of Out-of-Distribution Detection Performance on Autonomous Driving Datasets

2024-01-30 · Jens Henriksson, Christian Berger, Stig Ursing, Markus Borg

Safety measures need to be systemically investigated to what extent they evaluate the intended performance of Deep Neural Networks (DNNs) for critical applications. Due to a lack of verification methods for high-dimensio…

Autonomous DrivingOut-of-Distribution DetectionSemantic Segmentation

Frontier AI's Impact on the Cybersecurity Landscape

2025-04-07 · Wenbo Guo, Yujin Potter, Tianneng Shi, Zhun Wang 외

As frontier AI advances rapidly, understanding its impact on cybersecurity and inherent risks is essential to ensuring safe AI evolution (e.g., guiding risk mitigation and informing policymakers). While some studies revi…

Learning Material-Aware Hamiltonian Risk Fields for Safe Navigation

2026-05-07 · Aditya Sai Ellendula, Yi Wang, Chandrajit Bajaj arxiv

Risk-aware navigation should be selective: a policy should expose evasive degrees of freedom only when the local scene admits a lower-risk feasible maneuver, and suppress them when no safer alternative exists. We show th…