paper-with-me

홈 › Papers

On the Interpretability and Significance of Bias Metrics in Texts: a PMI-based Approach

2021-04-13 · Francisco Valentini, Germán Rosati, Damián Blasi, Diego Fernandez Slezak, Edgar Altszyler

In recent years, word embeddings have been widely used to measure biases in texts. Even if they have proven to be effective in detecting a wide variety of biases, metrics based on word embeddings lack transparency and interpretability. We analyze an alternative PMI-based metric to quantify biases in texts. It can be expressed as a function of conditional probabilities, which provides a simple interpretation in terms of word co-occurrences. We also prove that it can be approximated by an odds ratio, which allows estimating confidence intervals and statistical significance of textual biases. This approach produces similar results to metrics based on word embeddings when capturing gender gaps of the real world embedded in large corpora.

📄 PDF Abstract BibTeX arXiv:2104.06474

Code (1)

ftvalentini/biaspmi 공식 구현

Tasks

Word Embeddings

Similar Papers 제목 키워드 기반

On the interpretability and significance of bias metrics in texts: a PMI-based approach

2021-11-16 · ACL ARR November 2021 11 · Anonymous

In recent years, the use of word embeddings has become popular to measure the presence of biases in texts. Despite the fact that these measures have been proven to be effective in detecting a wide variety of biases, metr…

Word Embeddings

Data-Driven Analysis of Intersectional Bias in Image Classification: A Framework with Bias-Weighted Augmentation

2025-10-17 · Farjana Yesmin arxiv

Machine learning models trained on imbalanced datasets often exhibit intersectional biases-systematic errors arising from the interaction of multiple attributes such as object class and environmental conditions. This pap…

Image ClassificationData Augmentation

SHAMSUL: Systematic Holistic Analysis to investigate Medical Significance Utilizing Local interpretability methods in deep learning for chest radiography pathology prediction

2023-07-16 · Mahbub Ul Alam, Jaakko Hollmén, Jón Rúnar Baldvinsson, Rahim Rahmani

The interpretability of deep neural networks has become a subject of great interest within the medical and healthcare domain. This attention stems from concerns regarding transparency, legal and ethical considerations, a…

Transfer Learning

Towards Evaluating AI Systems for Moral Status Using Self-Reports

2023-11-14 · Ethan Perez, Robert Long

As AI systems become more advanced and widely deployed, there will likely be increasing debate over whether AI systems could have conscious experiences, desires, or other states of potential moral significance. It is imp…

Unbiased evaluation of ranking metrics reveals consistent performance in science and technology citation data

2020-01-15 · Shuqi Xu, Manuel Sebastian Mariani, Linyuan Lü, Matúš Medo

Despite the increasing use of citation-based metrics for research evaluation purposes, we do not know yet which metrics best deliver on their promise to gauge the significance of a scientific paper or a patent. We assess…

Information RetrievalRetrieval