paper-with-me

Papers

Building and Measuring Trust between Large Language Models

2025-08-20 · Maarten Buyl, Yousra Fettach, Guillaume Bied, Tijl De Bie arxiv

As large language models (LLMs) increasingly interact with each other, most notably in multi-agent setups, we may expect (and hope) that `trust' relationships develop between them, mirroring trust relationships between human colleagues, friends, or partners. Yet, though prior work has shown LLMs to be capable of identifying emotional connections and recognizing reciprocity in trust games, little remains known about (i) how different strategies to build trust compare, (ii) how such trust can be measured implicitly, and (iii) how this relates to explicit measures of trust. We study these questions by relating implicit measures of trust, i.e. susceptibility to persuasion and propensity to collaborate financially, with explicit measures of trust, i.e. a dyadic trust questionnaire well-established in psychology. We build trust in three ways: by building rapport dynamically, by starting from a prewritten script that evidences trust, and by adapting the LLMs' system prompt. Surprisingly, we find that the measures of explicit trust are either little or highly negatively correlated with implicit trust measures. These findings suggest that measuring trust between LLMs by asking their opinion may be deceiving. Instead, context-specific and implicit measures may be more informative in understanding how LLMs trust each other.

📄 PDF Abstract BibTeX arXiv:2508.15858

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Social Robot with Inner Speech for Dietary Guidance

2025-05-13 · Valerio Belcamino, Alessandro Carfì, Valeria Seidita, Fulvio Mastrogiovanni 외

We explore the use of inner speech as a mechanism to enhance transparency and trust in social robots for dietary advice. In humans, inner speech structures thought processes and decision-making; in robotics, it improves …

Computational EfficiencyDecision MakingNatural Language Understanding

Measuring and identifying factors of individuals' trust in Large Language Models

2025-02-28 · Edoardo Sebastiano De Duro, Giuseppe Alessandro Veltri, Hudson Golino, Massimo Stella

Large Language Models (LLMs) can engage in human-looking conversational exchanges. Although conversations can elicit trust between users and LLMs, scarce empirical research has examined trust formation in human-LLM conte…

Whether to trust: the ML leap of faith

2024-07-17 · Tory Frame, Julian Padget, George Stothart, Elizabeth Coulthard

Human trust is a prerequisite to trustworthy AI adoption, yet trust remains poorly understood. Trust is often described as an attitude, but attitudes cannot be reliably measured or managed. Additionally, humans frequentl…

Measuring Security Without Fooling Ourselves: Why Benchmarking Agents Is Hard

2026-05-21 · Sahar Abdelnabi, Chris Hicks, Konrad Rieck, Ahmad-Reza Sadeghi arxiv

The benchmarks used to evaluate AI agents in security-critical roles suffer from crucial weaknesses. Building on recent empirical evidence, we characterize three core challenges that undermine security evaluations: bench…

Measuring Consistency in Text-based Financial Forecasting Models

2023-05-15 · Linyi Yang, Yingpeng Ma, Yue Zhang

Financial forecasting has been an important and active area of machine learning research, as even the most modest advantage in predictive accuracy can be parlayed into significant financial gains. Recent advances in natu…