paper-with-me

Papers

When Trust Meets Truth: Trust-Truth Separability in LLM-as-Judge

2026-08-21 · Xin Sun, Di Wu, Yuchen Guo, Jiahuan Pei, Isao Echizen, Abdallah El Ali, Saku Sugawara arxiv

LLM-as-Judge systems can produce multi-dimensional evaluations, such as trustworthiness, reliability, and factuality, and these outputs are often interpreted as independent evidence. We test this assumption for a common pair of judgments: trust scoring and binary truth classification. On correctness-controlled QA, LLM judges align trust scores with truth verdicts more tightly than human behavioral reference, suggesting weaker separations between trust and truth judgment. We then apply stress tests by changing only source cues of identical QA between Human and AI. Source attribution shifts not only trust scores but also truth verdicts and logit-derived correct-side probabilities. Results show that current LLM-as-Judge protocols should not treat trust scores as independent evidence for truth judgments.

📄 PDF Abstract BibTeX arXiv:2608.21097

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Tell me the truth: A system to measure the trustworthiness of Large Language Models

2024-03-08 · Carlo Lipizzi

Large Language Models (LLM) have taken the front seat in most of the news since November 2022, when ChatGPT was introduced. After more than one year, one of the major reasons companies are resistant to adopting them is t…

Acoustic-Prosodic and Lexical Cues to Deception and Trust: Deciphering How People Detect Lies

2020-01-01 · TACL 2020 1 · Xi (Leslie) Chen, Sarah Ita Levitan, Michelle Levine, M 외

Humans rarely perform better than chance at lie detection. To better understand human perception of deception, we created a game framework, LieCatcher, to collect ratings of perceived deception using a large corpus of de…

Deception Detection

How model accuracy and explanation fidelity influence user trust

2019-07-26 · Andrea Papenmeier, Gwenn Englebienne, Christin Seifert

Machine learning systems have become popular in fields such as marketing, financing, or data mining. While they are highly accurate, complex machine learning systems pose challenges for engineers and users. Their inheren…

BIG-bench Machine LearningFairnessMarketing

It Takes Two to Lie: One to Lie, and One to Listen

2020-07-01 · ACL 2020 6 · Denis Peskov, Benny Cheng, Ahmed Elgohary, Joe Barrow 외

Trust is implicit in many online text conversations{---}striking up new friendships, or asking for tech support. But trust can be betrayed through deception. We study the language and dynamics of deception in the negotia…

Through the Looking-Glass: AI-Mediated Video Communication Reduces Interpersonal Trust and Confidence in Judgments

2026-03-19 · Nelson Navajas Fernández, Jeffrey T. Hancock, Maurice Jakesch arxiv

AI-based tools that mediate, enhance or generate parts of video communication may interfere with how people evaluate trustworthiness and credibility. In two preregistered online experiments (N = 2,000), we examined wheth…