paper-with-me

홈 › Papers

Exposing Bias in Online Communities through Large-Scale Language Models

2023-06-04 · Celine Wald, Lukas Pfahler

Progress in natural language generation research has been shaped by the ever-growing size of language models. While large language models pre-trained on web data can generate human-sounding text, they also reproduce social biases and contribute to the propagation of harmful stereotypes. This work utilises the flaw of bias in language models to explore the biases of six different online communities. In order to get an insight into the communities' viewpoints, we fine-tune GPT-Neo 1.3B with six social media datasets. The bias of the resulting models is evaluated by prompting the models with different demographics and comparing the sentiment and toxicity values of these generations. Together, these methods reveal that bias differs in type and intensity for the various models. This work not only affirms how easily bias is absorbed from training data but also presents a scalable method to identify and compare the bias of different datasets or communities. Additionally, the examples generated for this work demonstrate the limitations of using automated sentiment and toxicity classifiers in bias research.

📄 PDF Abstract BibTeX arXiv:2306.02294

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Methods 이 논문이 사용한 방법론

GPT-Neo An implementation of model & data parallel GPT3-like models using the mesh-tensorflow…

Similar Papers 제목 키워드 기반

Generating Ethnographic Models from Communities' Online Data

2020-07-01 · WS 2020 7 · Tomek Strzalkowski, Anna Newheiser, Nathan Kemper, Ning Sa 외

In this paper we describe computational ethnography study to demonstrate how machine learning techniques can be utilized to exploit bias resident in language data produced by communities with online presence. Specificall…

Discovering and Interpreting Biased Concepts in Online Communities

2020-10-27 · Xavier Ferrer-Aran, Tom van Nuenen, Natalia Criado, Jose M. Such

Language carries implicit human biases, functioning both as a reflection and a perpetuation of stereotypes that people carry with them. Recently, ML-based NLP methods such as word embeddings have been shown to learn such…

Cultural Vocal Bursts Intensity PredictionWord Embeddings

Comprehensive dataset of user-submitted articles with ideological and extreme bias from Reddit

2024-08-12 · Data in Brief 2024 8 · Kamalakkannan Ravi, Adan Ernesto Vela

Our study aims to collect data to understand ideological and extreme bias in text articles shared across various online communities, particularly focusing on the language used in subreddits associated with extremism and …

ArticlesHoldout SetNews ClassificationNews Recommendation+3

Discovering and Categorising Language Biases in Reddit

2020-08-06 · Xavier Ferrer, Tom van Nuenen, Jose M. Such, Natalia Criado

We present a data-driven approach using word embeddings to discover and categorise language biases on the discussion platform Reddit. As spaces for isolated user communities, platforms such as Reddit are increasingly con…

Word Embeddings

From Perils to Possibilities: Understanding how Human (and AI) Biases affect Online Fora

2024-03-21 · Virginia Morini, Valentina Pansanella, Katherine Abramski, Erica Cau 외

Social media platforms are online fora where users engage in discussions, share content, and build connections. This review explores the dynamics of social interactions, user-generated contents, and biases within the con…

Misinformation