paper-with-me

홈 › Papers

Measuring Gender Bias in Job Title Matching for Grammatical Gender Languages

2025-09-17 · Laura García-Sardiña, Hermenegildo Fabregat, Daniel Deniz, Rabih Zbib arxiv

This work sets the ground for studying how explicit grammatical gender assignment in job titles can affect the results of automatic job ranking systems. We propose the usage of metrics for ranking comparison controlling for gender to evaluate gender bias in job title ranking systems, in particular RBO (Rank-Biased Overlap). We generate and share test sets for a job title matching task in four grammatical gender languages, including occupations in masculine and feminine form and annotated by gender and matching relevance. We use the new test sets and the proposed methodology to evaluate the gender bias of several out-of-the-box multilingual models to set as baselines, showing that all of them exhibit varying degrees of gender bias.

📄 PDF Abstract BibTeX arXiv:2509.13803

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Measuring Gender Bias in Word Embeddings of Gendered Languages Requires Disentangling Grammatical Gender Signals

2022-06-03 · Shiva Omrani Sabbaghi, Aylin Caliskan

Does the grammatical gender of a language interfere when measuring the semantic gender information captured by its word embeddings? A number of anomalous gender bias measurements in the embeddings of gendered languages s…

Word Embeddings

ProText: A benchmark dataset for measuring (mis)gendering in long-form texts

2026-03-29 · Hadas Kotek, Margit Bowler, Patrick Sonnenberg, Yu'an Yang arxiv

We introduce ProText, a dataset for measuring gendering and misgendering in stylistically diverse long-form English texts. ProText spans three dimensions: Theme nouns (names, occupations, titles, kinship terms), Theme ca…

Evaluating Gender Bias Transfer from Film Data

2022-07-01 · NAACL (GeBNLP) 2022 7 · Amanda Bertsch, Ashley Oh, Sanika Natu, Swetha Gangu 외

Films are a rich source of data for natural language processing. OpenSubtitles (Lison and Tiedemann, 2016) is a popular movie script dataset, used for training models for tasks such as machine translation and dialogue ge…

Dialogue GenerationMachine TranslationSentenceSentence Embedding+2

Measuring and Mitigating Name Biases in Neural Machine Translation

2022-05-01 · ACL 2022 5 · Jun Wang, Benjamin Rubinstein, Trevor Cohn

Neural Machine Translation (NMT) systems exhibit problematic biases, such as stereotypical gender bias in the translation of occupation terms into languages with grammatical gender. In this paper we describe a new source…

Data AugmentationMachine TranslationNMTTranslation

Evaluating Gender Bias in Hindi-English Machine Translation

2021-06-16 · ACL (GeBNLP) 2021 8 · Gauri Gupta, Krithika Ramesh, Sanjay Singh

With language models being deployed increasingly in the real world, it is essential to address the issue of the fairness of their outputs. The word embedding representations of these language models often implicitly draw…

FairnessMachine TranslationSentenceTranslation