paper-with-me

Papers

Rethinking LLM Bias Probing Using Lessons from the Social Sciences

2025-02-28 · Kirsten N. Morehouse, Siddharth Swaroop, Weiwei Pan

The proliferation of LLM bias probes introduces three significant challenges: (1) we lack principled criteria for choosing appropriate probes, (2) we lack a system for reconciling conflicting results across probes, and (3) we lack formal frameworks for reasoning about when (and why) probe results will generalize to real user behavior. We address these challenges by systematizing LLM social bias probing using actionable insights from social sciences. We then introduce EcoLevels - a framework that helps (a) determine appropriate bias probes, (b) reconcile conflicting findings across probes, and (c) generate predictions about bias generalization. Overall, we ground our analysis in social science research because many LLM probes are direct applications of human probes, and these fields have faced similar challenges when studying social bias in humans. Based on our work, we suggest how the next generation of LLM bias probing can (and should) benefit from decades of social science research.

📄 PDF Abstract BibTeX arXiv:2503.00093

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Return of Pseudosciences in Artificial Intelligence: Have Machine Learning and Deep Learning Forgotten Lessons from Statistics and History?

2024-11-27 · Jérémie Sublime

In today's world, AI programs powered by Machine Learning are ubiquitous, and have achieved seemingly exceptional performance across a broad range of tasks, from medical diagnosis and credit rating in banking, to theft d…

Medical Diagnosis

Replication Markets: Results, Lessons, Challenges and Opportunities in AI Replication

2020-05-10 · Yang Liu, Michael Gordon, Juntao Wang, Michael Bishop 외

The last decade saw the emergence of systematic large-scale replication projects in the social and behavioral sciences, (Camerer et al., 2016, 2018; Ebersole et al., 2016; Klein et al., 2014, 2018; Collaboration, 2015). …

SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples

2023-11-30 · CVPR 2024 1 · Phillip Howard, Avinash Madasu, Tiep Le, Gustavo Lujan Moreno 외

While vision-language models (VLMs) have achieved remarkable performance improvements recently, there is growing evidence that these models also posses harmful biases with respect to social attributes such as gender and …

counterfactual

Rethinking Model Evaluation as Narrowing the Socio-Technical Gap

2023-06-01 · Q. Vera Liao, Ziang Xiao

The recent development of generative large language models (LLMs) poses new challenges for model evaluation that the research community and industry have been grappling with. While the versatile capabilities of these mod…

Explainable Artificial Intelligence (XAI)nlg evaluationvalid

Probing Intersectional Biases in Vision-Language Models with Counterfactual Examples

2023-10-04 · Phillip Howard, Avinash Madasu, Tiep Le, Gustavo Lujan Moreno 외

While vision-language models (VLMs) have achieved remarkable performance improvements recently, there is growing evidence that these models also posses harmful biases with respect to social attributes such as gender and …

counterfactual