paper-with-me

홈 › Papers

Detecting Natural Language Biases with Prompt-based Learning

2023-09-11 · Md Abdul Aowal, Maliha T Islam, Priyanka Mary Mammen, Sandesh Shetty

In this project, we want to explore the newly emerging field of prompt engineering and apply it to the downstream task of detecting LM biases. More concretely, we explore how to design prompts that can indicate 4 different types of biases: (1) gender, (2) race, (3) sexual orientation, and (4) religion-based. Within our project, we experiment with different manually crafted prompts that can draw out the subtle biases that may be present in the language model. We apply these prompts to multiple variations of popular and well-recognized models: BERT, RoBERTa, and T5 to evaluate their biases. We provide a comparative analysis of these models and assess them using a two-fold method: use human judgment to decide whether model predictions are biased and utilize model-level judgment (through further prompts) to understand if a model can self-diagnose the biases of its own prediction.

📄 PDF Abstract BibTeX arXiv:2309.05227

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingPrompt Engineering

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Residual Connection 설명 없음
Adam 설명 없음
Weight Decay 설명 없음

Similar Papers 제목 키워드 기반

AAVENUE: Detecting LLM Biases on NLU Tasks in AAVE via a Novel Benchmark

2024-08-27 · Abhay Gupta, Philip Meng, Ece Yurtseven, Sean O'Brien 외

Detecting biases in natural language understanding (NLU) for African American Vernacular English (AAVE) is crucial to developing inclusive natural language processing (NLP) systems. To address dialect-induced performance…

Language ModelingLanguage ModellingLarge Language ModelNatural Language Understanding

Using Natural Sentence Prompts for Understanding Biases in Language Models

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Evaluation of biases in language models is often limited to synthetically generated datasets. This dependence traces back to the need of prompt-style dataset to trigger specific behaviors of language models. In this pape…

Sentence

Using Natural Sentence Prompts for Understanding Biases in Language Models

2022-07-01 · NAACL 2022 7 · Sarah Alnegheimish, Alicia Guo, Yi Sun

Evaluation of biases in language models is often limited to synthetically generated datasets. This dependence traces back to the need of prompt-style dataset to trigger specific behaviors of language models. In this pape…

Sentence

Using Natural Sentences for Understanding Biases in Language Models

2022-05-12 · Sarah Alnegheimish, Alicia Guo, Yi Sun

Evaluation of biases in language models is often limited to synthetically generated datasets. This dependence traces back to the need for a prompt-style dataset to trigger specific behaviors of language models. In this p…

Sentence

Few-shot Instruction Prompts for Pretrained Language Models to Detect Social Biases

2021-12-15 · Shrimai Prabhumoye, Rafal Kocielnik, Mohammad Shoeybi, Anima Anandkumar 외

Detecting social bias in text is challenging due to nuance, subjectivity, and difficulty in obtaining good quality labeled datasets at scale, especially given the evolving nature of social biases and society. To address …