paper-with-me

홈 › Papers

Inducing anxiety in large language models can induce bias

2023-04-21 · Julian Coda-Forno, Kristin Witte, Akshay K. Jagadish, Marcel Binz, Zeynep Akata, Eric Schulz

Large language models (LLMs) are transforming research on machine learning while galvanizing public debates. Understanding not only when these models work well and succeed but also why they fail and misbehave is of great societal relevance. We propose to turn the lens of psychiatry, a framework used to describe and modify maladaptive behavior, to the outputs produced by these models. We focus on twelve established LLMs and subject them to a questionnaire commonly used in psychiatry. Our results show that six of the latest LLMs respond robustly to the anxiety questionnaire, producing comparable anxiety scores to humans. Moreover, the LLMs' responses can be predictably changed by using anxiety-inducing prompts. Anxiety-induction not only influences LLMs' scores on an anxiety questionnaire but also influences their behavior in a previously-established benchmark measuring biases such as racism and ageism. Importantly, greater anxiety-inducing text leads to stronger increases in biases, suggesting that how anxiously a prompt is communicated to large language models has a strong influence on their behavior in applied settings. These results demonstrate the usefulness of methods taken from psychiatry for studying the capable algorithms to which we increasingly delegate authority and autonomy.

📄 PDF Abstract BibTeX arXiv:2304.11111

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingPrompt Engineering

Methods 이 논문이 사용한 방법론

{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Multi-Head Attention 설명 없음
Attention 설명 없음
fail 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Adam 설명 없음

Similar Papers 제목 키워드 기반

Towards Controllable Biases in Language Generation

2020-05-01 · Findings of the Association for Computational Linguistics 2020 · Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, Nanyun Peng

We present a general approach towards controllable societal biases in natural language generation (NLG). Building upon the idea of adversarial triggers, we develop a method to induce societal biases in generated text whe…

Dialogue GenerationText Generation

From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data

2025-05-24 · Ugur Kursuncu, Trilok Padhi, Gaurav Sinha, Abdulkadir Erol 외

The growing demand for accessible mental health support, compounded by workforce shortages and logistical barriers, has led to increased interest in utilizing Large Language Models (LLMs) for scalable and real-time assis…

Supportive psychotherapy on insomnia induced by COVID-19; Evaluation of patients and hospital staff

2023-11-16 · Atieh Sadeghniiat-Haghighi, Arezu Najafi, Khosro Sadeghniiat Haghighi, Arghavan Shafiee-Aghdam 외

Introduction: The global COVID-19 pandemic has heightened stress, anxiety, and sadness, leading to increased rates of insomnia (6-10%). This study explores the effectiveness of supportive psychotherapy, specifically Cogn…

Sleep Quality

Who Trains Matters: Federated Learning under Enrollment and Participation Selection Biases

2026-04-29 · Gota Morishita arxiv

Federated learning (FL) trains a shared model from updates contributed by distributed clients, often implicitly assuming that contributing clients are representative of the target population. In practice, this representa…

Federated Learning

HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models

2025-03-17 · Xinyan Jiang, Hang Ye, Yongxin Zhu, Xiaoying Zheng 외

Large Language Models (LLMs) often generate hallucinations, producing outputs that are contextually inaccurate or factually incorrect. We introduce HICD, a novel method designed to induce hallucinations for contrastive d…

HallucinationQuestion AnsweringReading Comprehension