paper-with-me

Papers

Adversarial Negotiation Dynamics in Generative Language Models

2024-12-29 · Arinbjörn Kolbeinsson, Benedikt Kolbeinsson

Generative language models are increasingly used for contract drafting and enhancement, creating a scenario where competing parties deploy different language models against each other. This introduces not only a game-theory challenge but also significant concerns related to AI safety and security, as the language model employed by the opposing party can be unknown. These competitive interactions can be seen as adversarial testing grounds, where models are effectively red-teamed to expose vulnerabilities such as generating biased, harmful or legally problematic text. Despite the importance of these challenges, the competitive robustness and safety of these models in adversarial settings remain poorly understood. In this small study, we approach this problem by evaluating the performance and vulnerabilities of major open-source language models in head-to-head competitions, simulating real-world contract negotiations. We further explore how these adversarial interactions can reveal potential risks, informing the development of more secure and reliable models. Our findings contribute to the growing body of research on AI safety, offering insights into model selection and optimisation in competitive legal contexts and providing actionable strategies for mitigating risks.

📄 PDF Abstract BibTeX arXiv:2501.00069

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingModel Selection

Similar Papers 제목 키워드 기반

Behavioral Privacy Leakage in Agentic Negotiation: Formalizing and Mitigating Inference Attacks via Randomized Policies

2026-07-07 · Barkha Rani hf

Autonomous negotiation agents are increasingly deployed in high-stakes settings such as insurance and procurement. While cryptographic techniques protect explicitly disclosed constraint values, they fail to address a sub…

Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation

2023-09-29 · Sahar Abdelnabi, Amr Gomaa, Sarath Sivaprasad, Lea Schönherr 외

There is an growing interest in using Large Language Models (LLMs) in multi-agent systems to tackle interactive real-world tasks that require effective collaboration and assessing complex situations. Yet, we still have a…

Decision Making

EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation

2026-05-26 · Yunbo Long, Haolang Zhao, Lukas Beckenbauer, Liming Xu 외 arxiv

Post-trained LLMs are often optimized to align responses with human preferences, making them safe, polite, and conversationally appropriate. In adversarial negotiation, however, this alignment can become a vulnerability:…

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

2025-09-04 · Yunbo Long, Liming Xu, Lukas Beckenbauer, Yuhan Liu 외 arxiv

Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in \textit{complex}, \textit{multi-turn} negotiations, opening new avenues for agentic AI. Howev…

Reinforcement Learning

Advancing AI Negotiations: New Theory and Evidence from a Large-Scale Autonomous Negotiations Competition

2025-03-09 · Michelle Vaccaro, Michael Caoson, Harang Ju, Sinan Aral 외

Despite the rapid proliferation of artificial intelligence (AI) negotiation agents, there has been limited integration of computer science research and established negotiation theory to develop new theories of AI negotia…

Large Language Model