paper-with-me

Papers

ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science

2024-08-04 · Robert Wolfe, Alexis Hiniker, Bill Howe

This research introduces the Multilevel Embedding Association Test (ML-EAT), a method designed for interpretable and transparent measurement of intrinsic bias in language technologies. The ML-EAT addresses issues of ambiguity and difficulty in interpreting the traditional EAT measurement by quantifying bias at three levels of increasing granularity: the differential association between two target concepts with two attribute concepts; the individual effect size of each target concept with two attribute concepts; and the association between each individual target concept and each individual attribute concept. Using the ML-EAT, this research defines a taxonomy of EAT patterns describing the nine possible outcomes of an embedding association test, each of which is associated with a unique EAT-Map, a novel four-quadrant visualization for interpreting the ML-EAT. Empirical analysis of static and diachronic word embeddings, GPT-2 language models, and a CLIP language-and-image model shows that EAT patterns add otherwise unobservable information about the component biases that make up an EAT; reveal the effects of prompting in zero-shot models; and can also identify situations when cosine similarity is an ineffective metric, rendering an EAT unreliable. Our work contributes a method for rendering bias more observable and interpretable, improving the transparency of computational investigations into human minds and societies.

📄 PDF Abstract BibTeX arXiv:2408.01966

Code (1)

wolferobert3/ml-eat 공식 구현 pytorch

Tasks

AttributeDiachronic Word EmbeddingsWord Embeddings

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Residual Connection 설명 없음
Multi-Head Attention 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Weight Decay 설명 없음

Similar Papers 제목 키워드 기반

A Transparent Framework for Evaluating Unintended Demographic Bias in Word Embeddings

2019-07-01 · ACL 2019 7 · Chris Sweeney, Maryam Najafian

Word embedding models have gained a lot of traction in the Natural Language Processing community, however, they suffer from unintended demographic biases. Most approaches to evaluate these biases rely on vector space bas…

FairnessWord Embeddings

Interpretable EEG-to-Image Generation with Semantic Prompts

2025-07-09 · Arshak Rezvani, Ali Akbari, Kosar Sanjar Arani, Maryam Mirian 외 arxiv

Decoding visual experience from brain signals offers exciting possibilities for neuroscience and interpretable AI. While EEG is accessible and temporally precise, its limitations in spatial detail hinder image reconstruc…

Image ReconstructionContrastive LearningImage Generation

Interpretable Image Recognition by Constructing Transparent Embedding Space

2021-01-01 · ICCV 2021 10 · Jiaqi Wang, Huafeng Liu, Xinyue Wang, Liping Jing

Humans usually explain their reasoning (e.g. classification) by dissecting the image and pointing out the evidence from these parts to the concepts in their minds. Inspired by this cognitive process, several part-lev…

CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association

2026-06-26 · Zuoou Li, Wenlong Zhao, Kelly Yu, Weitong Zhang 외 arxiv

Identifying robust associations between cardiac imaging phenotypes and clinical diseases is fundamental to population-scale cardiovascular research and reliable risk stratification. However, current phenome-wide associat…

Word-Centered Semantic Graphs for Interpretable Diachronic Sense Tracking

2026-01-29 · Imene Kolli, Kai-Robin Lange, Jonas Rieger, Carsten Jentsch arxiv

We propose an interpretable, graph-based framework for analyzing semantic shift in diachronic corpora. For each target word and time slice, we induce a word-centered semantic network that integrates distributional simila…