paper-with-me

Papers

Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation

2024-06-02 · Bar Iluz, Yanai Elazar, Asaf Yehudai, Gabriel Stanovsky

Most works on gender bias focus on intrinsic bias -- removing traces of information about a protected group from the model's internal representation. However, these works are often disconnected from the impact of such debiasing on downstream applications, which is the main motivation for debiasing in the first place. In this work, we systematically test how methods for intrinsic debiasing affect neural machine translation models, by measuring the extrinsic bias of such systems under different design choices. We highlight three challenges and mismatches between the debiasing techniques and their end-goal usage, including the choice of embeddings to debias, the mismatch between words and sub-word tokens debiasing, and the effect on different target languages. We find that these considerations have a significant impact on downstream performance and the success of debiasing.

📄 PDF Abstract BibTeX arXiv:2406.00787

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing

2024-01-16 · Masahiro Kaneko, Danushka Bollegala, Timothy Baldwin

The output tendencies of Pre-trained Language Models (PLM) vary markedly before and after Fine-Tuning (FT) due to the updates to the model parameters. These divergences in output tendencies result in a gap in the social …

In-Context Learning

What Changed? Investigating Debiasing Methods using Causal Mediation Analysis

2022-06-01 · NAACL (GeBNLP) 2022 7 · Sullam Jeoung, Jana Diesner

Previous work has examined how debiasing language models affect downstream tasks, specifically, how debiasing techniques influence task performance and whether debiased models also make impartial predictions in downstrea…

How Gender Debiasing Affects Internal Model Representations, and Why It Matters

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Common studies of gender bias in NLP focus either on extrinsic bias measured by model performance on a downstream task or on intrinsic bias found in models' internal representations. However, the relationship between ext…

Debiasing Isn’t Enough! – on the Effectiveness of Debiasing MLMs and Their Social Biases in Downstream Tasks

2022-10-01 · COLING 2022 10 · Masahiro Kaneko, Danushka Bollegala, Naoaki Okazaki

We study the relationship between task-agnostic intrinsic and task-specific extrinsic social bias evaluation measures for MLMs, and find that there exists only a weak correlation between these two types of evaluation mea…

How Gender Debiasing Affects Internal Model Representations, and Why It Matters

2022-04-14 · NAACL 2022 7 · Hadas Orgad, Seraphina Goldfarb-Tarrant, Yonatan Belinkov

Common studies of gender bias in NLP focus either on extrinsic bias measured by model performance on a downstream task or on intrinsic bias found in models' internal representations. However, the relationship between ext…