paper-with-me

홈 › Papers

Quantifying Risk Propensities of Large Language Models: Ethical Focus and Bias Detection through Role-Play

2024-10-26 · Yifan Zeng, Liang Kairong, Fangzhou Dong, Peijia Zheng

As Large Language Models (LLMs) become more prevalent, concerns about their safety, ethics, and potential biases have risen. Systematically evaluating LLMs' risk decision-making tendencies and attitudes, particularly in the ethical domain, has become crucial. This study innovatively applies the Domain-Specific Risk-Taking (DOSPERT) scale from cognitive science to LLMs and proposes a novel Ethical Decision-Making Risk Attitude Scale (EDRAS) to assess LLMs' ethical risk attitudes in depth. We further propose a novel approach integrating risk scales and role-playing to quantitatively evaluate systematic biases in LLMs. Through systematic evaluation and analysis of multiple mainstream LLMs, we assessed the "risk personalities" of LLMs across multiple domains, with a particular focus on the ethical domain, and revealed and quantified LLMs' systematic biases towards different groups. This research helps understand LLMs' risk decision-making and ensure their safe and reliable application. Our approach provides a tool for identifying and mitigating biases, contributing to fairer and more trustworthy AI systems. The code and data are available.

📄 PDF Abstract BibTeX arXiv:2411.08884

Code (0)

등록된 구현이 없습니다.

Tasks

Bias DetectionDecision MakingEthics

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Sycophancy in Large Language Models: Causes and Mitigations

2024-11-22 · Lars Malmqvist

Large language models (LLMs) have demonstrated remarkable capabilities across a wide range of natural language processing tasks. However, their tendency to exhibit sycophantic behavior - excessively agreeing with or flat…

Hallucination

When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing

2025-03-30 · Haein Kong, Seonghyeon Moon

Large language models (LLMs) have been actively applied in the mental health field. Recent research shows the promise of LLMs in applying psychotherapy, especially motivational interviewing (MI). However, there is a lack…

EthicsResponse Generation

AI Act and Large Language Models (LLMs): When critical issues and privacy impact require human and ethical oversight

2024-03-31 · Nicola Fabiano

The imposing evolution of artificial intelligence systems and, specifically, of Large Language Models (LLM) makes it necessary to carry out assessments of their level of risk and the impact they may have in the area of p…

Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey

2024-06-08 · Chengyuan Deng, Yiqun Duan, Xin Jin, Heng Chang 외

Large Language Models (LLMs) have achieved unparalleled success across diverse language modeling tasks in recent years. However, this progress has also intensified ethical concerns, impacting the deployment of LLMs in ev…

EthicsLanguage ModelingLanguage ModellingSurvey

Ethical-Advice Taker: Do Language Models Understand Natural Language Interventions?

2021-06-02 · Findings (ACL) 2021 8 · Jieyu Zhao, Daniel Khashabi, Tushar Khot, Ashish Sabharwal 외

Is it possible to use natural language to intervene in a model's behavior and alter its prediction in a desired way? We investigate the effectiveness of natural language interventions for reading-comprehension systems, s…

EthicsFew-Shot LearningQuestion AnsweringReading Comprehension