paper-with-me

홈 › Papers

NLPositionality: Characterizing Design Biases of Datasets and Models

2023-06-02 · Sebastin Santy, Jenny T. Liang, Ronan Le Bras, Katharina Reinecke, Maarten Sap

Design biases in NLP systems, such as performance differences for different populations, often stem from their creator's positionality, i.e., views and lived experiences shaped by identity and background. Despite the prevalence and risks of design biases, they are hard to quantify because researcher, system, and dataset positionality is often unobserved. We introduce NLPositionality, a framework for characterizing design biases and quantifying the positionality of NLP datasets and models. Our framework continuously collects annotations from a diverse pool of volunteer participants on LabintheWild, and statistically quantifies alignment with dataset labels and model predictions. We apply NLPositionality to existing datasets and models for two tasks -- social acceptability and hate speech detection. To date, we have collected 16,299 annotations in over a year for 600 instances from 1,096 annotators across 87 countries. We find that datasets and models align predominantly with Western, White, college-educated, and younger populations. Additionally, certain groups, such as non-binary people and non-native English speakers, are further marginalized by datasets and models as they rank least in alignment across all tasks. Finally, we draw from prior literature to discuss how researchers can examine their own positionality and that of their datasets and models, opening the door for more inclusive NLP systems.

📄 PDF Abstract BibTeX arXiv:2306.01943

Code (1)

liang-jenny/nlpositionality 공식 구현

Tasks

Hate Speech Detection

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Automatically Characterizing Targeted Information Operations Through Biases Present in Discourse on Twitter

2020-04-18 · Autumn Toney, Akshat Pandey, Wei Guo, David Broniatowski 외

This paper considers the problem of automatically characterizing overall attitudes and biases that may be associated with emerging information operations via artificial intelligence. Accurate analysis of these emerging t…

TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models

2025-11-30 · Maya Varma, Jean-Benoit Delbrouck, Sophie Ostmeier, Akshay Chaudhari 외 arxiv

Vision-language models (VLMs) have made great strides in addressing temporal understanding tasks, which involve characterizing visual changes across a sequence of images. However, recent works have suggested that when ma…

Relational inductive biases on attention mechanisms

2025-07-05 · Víctor Mijangos, Ximena Gutierrez-Vasques, Verónica E. Arriola, Ulises Rodríguez-Domínguez 외 arxiv

Inductive learning aims to construct general models from specific examples, guided by biases that influence hypothesis selection and determine generalization capacity. In this work, we focus on characterizing the relatio…

Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis

2019-08-01 · WS 2019 8 · Scott Friedman, Sonja Schmer-Galunder, Anthony Chen, Jeffrey Rye

Modern models for common NLP tasks often employ machine learning techniques and train on journalistic, social media, or other culturally-derived text. These have recently been scrutinized for racial and gender biases, ro…

Cultural Vocal Bursts Intensity PredictionWord Embeddings

Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis

2026-01-23 · Scott Friedman, Sonja Schmer-Galunder, Anthony Chen, Jeffrey Rye arxiv

Modern models for common NLP tasks often employ machine learning techniques and train on journalistic, social media, or other culturally-derived text. These have recently been scrutinized for racial and gender biases, ro…