paper-with-me

Papers

CORGI-PM: A Chinese Corpus For Gender Bias Probing and Mitigation

2023-01-01 · Ge Zhang, Yizhi Li, Yaoyao Wu, Linyuan Zhang, Chenghua Lin, Jiayi Geng, Shi Wang, Jie Fu

As natural language processing (NLP) for gender bias becomes a significant interdisciplinary topic, the prevalent data-driven techniques such as large-scale language models suffer from data inadequacy and biased corpus, especially for languages with insufficient resources such as Chinese. To this end, we propose a Chinese cOrpus foR Gender bIas Probing and Mitigation CORGI-PM, which contains 32.9k sentences with high-quality labels derived by following an annotation scheme specifically developed for gender bias in the Chinese context. Moreover, we address three challenges for automatic textual gender bias mitigation, which requires the models to detect, classify, and mitigate textual gender bias. We also conduct experiments with state-of-the-art language models to provide baselines. To our best knowledge, CORGI-PM is the first sentence-level Chinese corpus for gender bias probing and mitigation.

📄 PDF Abstract BibTeX arXiv:2301.00395

Code (1)

yizhilll/corgi-pm 공식 구현 pytorch

Tasks

Sentence

Similar Papers 제목 키워드 기반

Gender Bias in Contextualized Word Embeddings

2019-04-05 · NAACL 2019 6 · Jieyu Zhao, Tianlu Wang, Mark Yatskar, Ryan Cotterell 외

In this paper, we quantify, analyze and mitigate gender bias exhibited in ELMo's contextualized word vectors. First, we conduct several intrinsic analyses and find that (1) training data for ELMo contains significantly m…

Word Embeddings

Contextual Tokenization for Graph Inverted Indices

2025-10-26 · Pritish Chakraborty, Indradyumna Roy, Soumen Chakrabarti, Abir De arxiv

Retrieving graphs from a large corpus, that contain a subgraph isomorphic to a given query graph, is a core operation in many real-world applications. While recent multi-vector graph representations and scores based on s…

Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups

2025-10-08 · Geng Liu, Feng Li, Junjie Mu, Mengxiao Zhu 외 arxiv

Large language models (LLMs) are increasingly deployed in user-facing applications, raising concerns that they may reflect and amplify social biases. We investigate social identity biases in Chinese LLMs using Mandarin-s…

Gender Bias in Natural Language Processing Across Human Languages

2021-06-01 · NAACL (TrustNLP) 2021 6 · Abigail Matthews, Isabella Grasso, Christopher Mahoney, Yan Chen 외

Natural Language Processing (NLP) systems are at the heart of many critical automated decision-making systems making crucial recommendations about our future world. Gender bias in NLP has been well studied in English, bu…

Decision Making

Gender Bias Hidden Behind Chinese Word Embeddings: The Case of Chinese Adjectives

2021-06-01 · ACL (GeBNLP) 2021 8 · Meichun Jiao, Ziyang Luo

Gender bias in word embeddings gradually becomes a vivid research field in recent years. Most studies in this field aim at measurement and debiasing methods with English as the target language. This paper investigates ge…

Word Embeddings