paper-with-me

홈 › Papers

Linguistic Properties of Truthful Response

2023-05-25 · Bruce W. Lee, Benedict Florance Arockiaraj, Helen Jin

We investigate the phenomenon of an LLM's untruthful response using a large set of 220 handcrafted linguistic features. We focus on GPT-3 models and find that the linguistic profiles of responses are similar across model sizes. That is, how varying-sized LLMs respond to given prompts stays similar on the linguistic properties level. We expand upon this finding by training support vector machines that rely only upon the stylistic components of model responses to classify the truthfulness of statements. Though the dataset size limits our current findings, we show the possibility that truthfulness detection is possible without evaluating the content itself. But at the same time, the limited scope of our experiments must be taken into account in interpreting the results.

📄 PDF Abstract BibTeX arXiv:2305.15875

Code (1)

benedictflorance/truthfulqa_experiments

Methods 이 논문이 사용한 방법론

{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
Multi-Head Attention 설명 없음
Attention 설명 없음
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…

Similar Papers 제목 키워드 기반

Linguistic Cues to Deception and Perceived Deception in Interview Dialogues

2018-06-01 · NAACL 2018 6 · Sarah Ita Levitan, Angel Maredia, Julia Hirschberg

We explore deception detection in interview dialogues. We analyze a set of linguistic features in both truthful and deceptive responses to interview questions. We also study the perception of deception, identifying chara…

BIG-bench Machine LearningDeception DetectionGeneral Classification

When lies are mostly truthful: automated verbal deception detection for embedded lies

2025-01-13 · Riccardo Loconte, Bennett Kleinberg

Background: Verbal deception detection research relies on narratives and commonly assumes statements as truthful or deceptive. A more realistic perspective acknowledges that the veracity of statements exists on a continu…

Deception Detection

The Square Root Agreement Rule for Incentivizing Truthful Feedback on Online Platforms

2015-07-25 · Vijay Kamble, Nihar Shah, David Marn, Abhay Parekh 외

A major challenge in obtaining evaluations of products or services on e-commerce platforms is eliciting informative responses in the absence of verifiability. This paper proposes the Square Root Agreement Rule (SRA): a s…

TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space

2024-02-27 · Shaolei Zhang, Tian Yu, Yang Feng

Large Language Models (LLMs) sometimes suffer from producing hallucinations, especially LLMs may generate untruthful responses despite knowing the correct knowledge. Activating the truthfulness within LLM is the key to f…

Contrastive LearningHallucinationHallucination EvaluationLanguage Modelling+4

Acoustic-Prosodic and Lexical Cues to Deception and Trust: Deciphering How People Detect Lies

2020-01-01 · TACL 2020 1 · Xi (Leslie) Chen, Sarah Ita Levitan, Michelle Levine, M 외

Humans rarely perform better than chance at lie detection. To better understand human perception of deception, we created a game framework, LieCatcher, to collect ratings of perceived deception using a large corpus of de…

Deception Detection