paper-with-me

홈 › Papers

Demographic Biases and Gaps in the Perception of Sexism in Large Language Models

2025-08-25 · Judith Tavarez-Rodríguez, Fernando Sánchez-Vega, A. Pastor López-Monroy arxiv

The use of Large Language Models (LLMs) has proven to be a tool that could help in the automatic detection of sexism. Previous studies have shown that these models contain biases that do not accurately reflect reality, especially for minority groups. Despite various efforts to improve the detection of sexist content, this task remains a significant challenge due to its subjective nature and the biases present in automated models. We explore the capabilities of different LLMs to detect sexism in social media text using the EXIST 2024 tweet dataset. It includes annotations from six distinct profiles for each tweet, allowing us to evaluate to what extent LLMs can mimic these groups' perceptions in sexism detection. Additionally, we analyze the demographic biases present in the models and conduct a statistical analysis to identify which demographic characteristics (age, gender) contribute most effectively to this task. Our results show that, while LLMs can to some extent detect sexism when considering the overall opinion of populations, they do not accurately replicate the diversity of perceptions among different demographic groups. This highlights the need for better-calibrated models that account for the diversity of perspectives across different populations.

📄 PDF Abstract BibTeX arXiv:2508.18245

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Effects of Demographic Instructions on LLM Personas

2025-05-17 · Angel Felipe Magnossão de Paula, J. Shane Culpepper, Alistair Moffat, Sachin Pathiyan Cherumanal 외

Social media platforms must filter sexist content in compliance with governmental regulations. Current machine learning approaches can reliably detect sexism based on standardized definitions, but often neglect the subje…

Large scale analysis of gender bias and sexism in song lyrics

2022-08-03 · Lorenzo Betti, Carlo Abrate, Andreas Kaltenbrunner

We employ Natural Language Processing techniques to analyse 377808 English song lyrics from the "Two Million Song Database" corpus, focusing on the expression of sexism across five decades (1960-2010) and the measurement…

Cultural Vocal Bursts Intensity PredictionWord Embeddings

Detecting clinician implicit biases in diagnoses using proximal causal inference

2025-01-27 · Kara Liu, Russ Altman, Vasilis Syrgkanis

Clinical decisions to treat and diagnose patients are affected by implicit biases formed by racism, ableism, sexism, and other stereotypes. These biases reflect broader systemic discrimination in healthcare and risk marg…

AttributeCausal Inference

Sociolectal Analysis of Pretrained Language Models

2021-11-01 · EMNLP 2021 11 · Sheng Zhang, Xin Zhang, Weiming Zhang, Anders Søgaard

Using data from English cloze tests, in which subjects also self-reported their gender, age, education, and race, we examine performance differences of pretrained language models across demographic groups, defined by the…

Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study

2025-10-06 · Ayan Majumdar, Feihao Chen, Jinghui Li, Xiaozhen Wang arxiv

Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a regulatory need for data auditing and developing scalable bias-detection m…

Bias Detection