paper-with-me

홈 › Papers

CommunityLM: Probing Partisan Worldviews from Language Models

2022-09-15 · COLING 2022 10 · Hang Jiang, Doug Beeferman, Brandon Roy, Deb Roy

As political attitudes have diverged ideologically in the United States, political speech has diverged lingusitically. The ever-widening polarization between the US political parties is accelerated by an erosion of mutual understanding between them. We aim to make these communities more comprehensible to each other with a framework that probes community-specific responses to the same survey questions using community language models CommunityLM. In our framework we identify committed partisan members for each community on Twitter and fine-tune LMs on the tweets authored by them. We then assess the worldviews of the two groups using prompt-based probing of their corresponding LMs, with prompts that elicit opinions about public figures and groups surveyed by the American National Election Studies (ANES) 2020 Exploratory Testing Survey. We compare the responses generated by the LMs to the ANES survey results, and find a level of alignment that greatly exceeds several baseline methods. Our work aims to show that we can use community LMs to query the worldview of any group of people given a sufficiently large sample of their social media discussions or media diet.

📄 PDF Abstract BibTeX arXiv:2209.07065

Code (1)

hjian42/communitylm 공식 구현 pytorch

Tasks

Survey

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Reading Between the Tweets: Deciphering Ideological Stances of Interconnected Mixed-Ideology Communities

2024-02-02 · Zihao He, Ashwin Rao, Siyi Guo, Negar Mokhberian 외

Recent advances in NLP have improved our ability to understand the nuanced worldviews of online communities. Existing research focused on probing ideological stances treats liberals and conservatives as separate groups. …

The Neutral Mask: How RLHF Provides Shallow Alignment while Leaving Partisan Structure Intact in a Large Language Model

2026-06-08 · Wendy K. Tam arxiv

The ambition behind alignment training is to make large language models safe and useful. The primary mechanism, reinforcement learning from human feedback (RLHF), shapes the behavior of deployed language models by aligni…

Reinforcement Learning

PartisanLens: A Multilingual Dataset of Hyperpartisan and Conspiratorial Immigration Narratives in European Media

2026-01-07 · Michele Joshua Maggini, Paloma Piot, Anxo Pérez, Erik Bran Marino 외 arxiv

Detecting hyperpartisan narratives and Population Replacement Conspiracy Theories (PRCT) is essential to addressing the spread of misinformation. These complex narratives pose a significant threat, as hyperpartisanship d…

Avoiding Disparity Amplification under Different Worldviews

2018-08-26 · Samuel Yeom, Michael Carl Tschantz

We mathematically compare four competing definitions of group-level nondiscrimination: demographic parity, equalized odds, predictive parity, and calibration. Using the theoretical framework of Friedler et al., we study …

Fairness

Crossing the Aisle: Unveiling Partisan and Counter-Partisan Events in News Reporting

2023-10-28 · Kaijian Zou, Xinliang Frederick Zhang, Winston Wu, Nick Beauchamp 외

News media is expected to uphold unbiased reporting. Yet they may still affect public opinion by selectively including or omitting events that support or contradict their ideological positions. Prior work in NLP has only…

Articles