paper-with-me

홈 › Papers

Revealing Fine-Grained Values and Opinions in Large Language Models

2024-06-27 · Dustin Wright, Arnav Arora, Nadav Borenstein, Srishti Yadav, Serge Belongie, Isabelle Augenstein

Uncovering latent values and opinions embedded in large language models (LLMs) can help identify biases and mitigate potential harm. Recently, this has been approached by prompting LLMs with survey questions and quantifying the stances in the outputs towards morally and politically charged statements. However, the stances generated by LLMs can vary greatly depending on how they are prompted, and there are many ways to argue for or against a given position. In this work, we propose to address this by analysing a large and robust dataset of 156k LLM responses to the 62 propositions of the Political Compass Test (PCT) generated by 6 LLMs using 420 prompt variations. We perform coarse-grained analysis of their generated stances and fine-grained analysis of the plain text justifications for those stances. For fine-grained analysis, we propose to identify tropes in the responses: semantically similar phrases that are recurrent and consistent across different prompts, revealing natural patterns in the text that a given LLM is prone to produce. We find that demographic features added to prompts significantly affect outcomes on the PCT, reflecting bias, as well as disparities between the results of tests when eliciting closed-form vs. open domain responses. Additionally, patterns in the plain text rationales via tropes show that similar justifications are repeatedly generated across models and prompts even with disparate stances.

📄 PDF Abstract BibTeX arXiv:2406.19238

Code (1)

copenlu/llm-pct-tropes 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

PCT 설명 없음

Similar Papers 제목 키워드 기반

Fine-tuning language models to find agreement among humans with diverse preferences

2022-11-28 · Michiel A. Bakker, Martin J. Chadwick, Hannah R. Sheahan, Michael Henry Tessler 외

Recent work in large language modeling (LLMs) has used fine-tuning to align outputs with the preferences of a prototypical user. This work assumes that human preferences are static and homogeneous across individuals, so …

Language ModelingLanguage Modelling

From Values to Opinions: Predicting Human Behaviors and Stances Using Value-Injected Large Language Models

2023-10-27 · Dongjun Kang, Joonsuk Park, Yohan Jo, JinYeong Bak

Being able to predict people's opinions on issues and behaviors in realistic scenarios can be helpful in various domains, such as politics and marketing. However, conducting large-scale surveys like the European Social S…

MarketingQuestion Answering

Fine-grained Financial Opinion Mining: A Survey and Research Agenda

2020-05-05 · Chung-Chi Chen, Hen-Hsen Huang, Hsin-Hsi Chen

Opinion mining is a prevalent research issue in many domains. In the financial domain, however, it is still in the early stages. Most of the researches on this topic only focus on the coarse-grained market sentiment anal…

Opinion MiningPositionSentiment AnalysisSurvey

From the Token to the Review: A Hierarchical Multimodal approach to Opinion Mining

2019-08-29 · IJCNLP 2019 11 · Alexandre Garcia, Pierre Colombo, Slim Essid, Florence d'Alché-Buc 외

The task of predicting fine grained user opinion based on spontaneous spoken language is a key problem arising in the development of Computational Agents as well as in the development of social network based opinion mine…

Opinion Mining

Fine-grained German Sentiment Analysis on Social Media

2012-05-01 · LREC 2012 5 · Saeedeh Momtazi

Expressing opinions and emotions on social media becomes a frequent activity in daily life. People express their opinions about various targets via social media and they are also interested to know about other opinions o…

Sentiment Analysis