paper-with-me

Papers

Escalation Risks from Language Models in Military and Diplomatic Decision-Making

2024-01-07 · Juan-Pablo Rivera, Gabriel Mukobi, Anka Reuel, Max Lamparth, Chandler Smith, Jacquelyn Schneider

Governments are increasingly considering integrating autonomous AI agents in high-stakes military and foreign-policy decision-making, especially with the emergence of advanced generative AI models like GPT-4. Our work aims to scrutinize the behavior of multiple AI agents in simulated wargames, specifically focusing on their predilection to take escalatory actions that may exacerbate multilateral conflicts. Drawing on political science and international relations literature about escalation dynamics, we design a novel wargame simulation and scoring framework to assess the escalation risks of actions taken by these agents in different scenarios. Contrary to prior studies, our research provides both qualitative and quantitative insights and focuses on large language models (LLMs). We find that all five studied off-the-shelf LLMs show forms of escalation and difficult-to-predict escalation patterns. We observe that models tend to develop arms-race dynamics, leading to greater conflict, and in rare cases, even to the deployment of nuclear weapons. Qualitatively, we also collect the models' reported reasonings for chosen actions and observe worrying justifications based on deterrence and first-strike tactics. Given the high stakes of military and foreign-policy contexts, we recommend further examination and cautious consideration before deploying autonomous language model agents for strategic military or diplomatic decision-making.

📄 PDF Abstract BibTeX arXiv:2401.03408

Code (1)

jprivera44/EscalAItion 공식 구현 pytorch

Tasks

Decision MakingLanguage Modelling

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Adam 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Residual Connection 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…

Similar Papers 제목 키워드 기반

Critical Foreign Policy Decisions (CFPD)-Benchmark: Measuring Diplomatic Preferences in Large Language Models

2025-03-08 · Benjamin Jensen, Ian Reynolds, Yasir Atalan, Michael Garcia 외

As national security institutions increasingly integrate Artificial Intelligence (AI) into decision-making and content generation processes, understanding the inherent biases of large language models (LLMs) is crucial. T…

Humanitarian

Military AI Needs Technically-Informed Regulation to Safeguard AI Research and its Applications

2025-05-23 · Riley Simmons-Edler, Jean Dong, Paul Lushenko, Kanaka Rajan 외

Military weapon systems and command-and-control infrastructure augmented by artificial intelligence (AI) have seen rapid development and deployment in recent years. However, the sociotechnical impacts of AI on combat sys…

Human vs. Machine: Behavioral Differences Between Expert Humans and Language Models in Wargame Simulations

2024-03-06 · Max Lamparth, Anthony Corso, Jacob Ganz, Oriana Skylar Mastro 외

To some, the advent of artificial intelligence (AI) promises better decision-making and increased military effectiveness while reducing the influence of human error and emotions. However, there is still debate about how …

Decision Making

Resilience Through Escalation: A Graph-Based PACE Architecture for Satellite Threat Response

2025-06-25 · Anouar Boumeftah, Sarah McKenzie-Picot, Peter Klimas, Gunes Karabulut Kurt

Satellite systems increasingly face operational risks from jamming, cyberattacks, and electromagnetic disruptions. Traditional redundancy strategies often fail against dynamic, multi-vector threats. This paper introduces…

AI-Powered Autonomous Weapons Risk Geopolitical Instability and Threaten AI Research

2024-05-03 · Riley Simmons-Edler, Ryan Badman, Shayne Longpre, Kanaka Rajan

The recent embrace of machine learning (ML) in the development of autonomous weapons systems (AWS) creates serious risks to geopolitical stability and the free exchange of ideas in AI research. This topic has received co…