paper-with-me

Papers

Evaluating Bias in LLMs for Job-Resume Matching: Gender, Race, and Education

2025-03-24 · Hayate Iso, Pouya Pezeshkpour, Nikita Bhutani, Estevam Hruschka

Large Language Models (LLMs) offer the potential to automate hiring by matching job descriptions with candidate resumes, streamlining recruitment processes, and reducing operational costs. However, biases inherent in these models may lead to unfair hiring practices, reinforcing societal prejudices and undermining workplace diversity. This study examines the performance and fairness of LLMs in job-resume matching tasks within the English language and U.S. context. It evaluates how factors such as gender, race, and educational background influence model decisions, providing critical insights into the fairness and reliability of LLMs in HR applications. Our findings indicate that while recent models have reduced biases related to explicit attributes like gender and race, implicit biases concerning educational background remain significant. These results highlight the need for ongoing evaluation and the development of advanced bias mitigation strategies to ensure equitable hiring practices when using LLMs in industry settings.

📄 PDF Abstract BibTeX arXiv:2503.19182

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityFairness

Similar Papers 제목 키워드 기반

Are Emily and Greg Still More Employable than Lakisha and Jamal? Investigating Algorithmic Hiring Bias in the Era of ChatGPT

2023-10-08 · Akshaj Kumar Veldanda, Fabian Grob, Shailja Thakur, Hammond Pearce 외

Large Language Models (LLMs) such as GPT-3.5, Bard, and Claude exhibit applicability across numerous tasks. One domain of interest is their use in algorithmic hiring, specifically in matching resumes with job categories.…

Evaluation of Bias Towards Medical Professionals in Large Language Models

2024-06-30 · Xi Chen, Yang Xu, MingKe You, Li Wang 외

This study evaluates whether large language models (LLMs) exhibit biases towards medical professionals. Fictitious candidate resumes were created to control for identity factors while maintaining consistent qualification…

The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring

2024-05-07 · Lena Armstrong, Abbey Liu, Stephen MacNeil, Danaë Metaxa

Large language models (LLMs) are increasingly being introduced in workplace settings, with the goals of improving efficiency and fairness. However, concerns have arisen regarding these models' potential to reflect or exa…

Fairness

JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models

2024-06-17 · Ze Wang, Zekun Wu, Xin Guan, Michael Thaler 외

The use of Large Language Models (LLMs) in hiring has led to legislative actions to protect vulnerable demographic groups. This paper presents a novel framework for benchmarking hierarchical gender hiring bias in Large L…

BenchmarkingcounterfactualResume Scoring

FAIRE: Assessing Racial and Gender Bias in AI-Driven Resume Evaluations

2025-04-02 · Athena Wen, Tanush Patil, Ansh Saxena, Yicheng Fu 외

In an era where AI-driven hiring is transforming recruitment practices, concerns about fairness and bias have become increasingly important. To explore these issues, we introduce a benchmark, FAIRE (Fairness Assessment I…

Fairness