paper-with-me

홈 › Papers

Finding Concept-specific Biases in Form--Meaning Associations

2021-04-13 · NAACL 2021 4 · Tiago Pimentel, Brian Roark, Søren Wichmann, Ryan Cotterell, Damián Blasi

This work presents an information-theoretic operationalisation of cross-linguistic non-arbitrariness. It is not a new idea that there are small, cross-linguistic associations between the forms and meanings of words. For instance, it has been claimed (Blasi et al., 2016) that the word for "tongue" is more likely than chance to contain the phone [l]. By controlling for the influence of language family and geographic proximity within a very large concept-aligned, cross-lingual lexicon, we extend methods previously used to detect within language non-arbitrariness (Pimentel et al., 2019) to measure cross-linguistic associations. We find that there is a significant effect of non-arbitrariness, but it is unsurprisingly small (less than 0.5% on average according to our information-theoretic estimate). We also provide a concept-level analysis which shows that a quarter of the concepts considered in our work exhibit a significant level of cross-linguistic non-arbitrariness. In sum, the paper provides new methods to detect cross-linguistic associations at scale, and confirms their effects are minor.

📄 PDF Abstract BibTeX arXiv:2104.06325

Code (2)

rycolab/form-meaning-associations 공식 구현 pytorch
tpimentelms/form-meaning-associations 공식 구현 pytorch

Tasks

Form

Similar Papers 제목 키워드 기반

Vision-Language Models Performing Zero-Shot Tasks Exhibit Gender-based Disparities

2023-01-26 · Melissa Hall, Laura Gustafson, Aaron Adcock, Ishan Misra 외

We explore the extent to which zero-shot vision-language models exhibit gender bias for different vision tasks. Vision models traditionally required task-specific labels for representing concepts, as well as finetuning; …

image-classificationImage Classificationobject-detectionObject Detection+3

Discovering and Interpreting Biased Concepts in Online Communities

2020-10-27 · Xavier Ferrer-Aran, Tom van Nuenen, Natalia Criado, Jose M. Such

Language carries implicit human biases, functioning both as a reflection and a perpetuation of stereotypes that people carry with them. Recently, ML-based NLP methods such as word embeddings have been shown to learn such…

Cultural Vocal Bursts Intensity PredictionWord Embeddings

Breaking Down Bias: On The Limits of Generalizable Pruning Strategies

2025-02-11 · Sibo Ma, Alejandro Salinas, Peter Henderson, Julian Nyarko

We employ model pruning to examine how LLMs conceptualize racial biases, and whether a generalizable mitigation strategy for such biases appears feasible. Our analysis yields several novel insights. We find that pruning …

Decision Making

Meaningfully Debugging Model Mistakes using Conceptual Counterfactual Explanations

2021-06-24 · Abubakar Abid, Mert Yuksekgonul, James Zou

Understanding and explaining the mistakes made by trained models is critical to many machine learning objectives, such as improving robustness, addressing concept drift, and mitigating biases. However, this is often an a…

counterfactualmodel

A Simple, Yet Effective Approach to Finding Biases in Code Generation

2022-10-31 · Spyridon Mouselinos, Mateusz Malinowski, Henryk Michalewski

Recently, high-performing code generation systems based on large language models have surfaced. They are trained on massive corpora containing much more natural text than actual executable computer code. This work shows …

Causal Language ModelingCode GenerationLanguage ModelingLanguage Modelling+1