paper-with-me

홈 › Papers

Markedness in Visual Semantic AI

2022-05-23 · Robert Wolfe, Aylin Caliskan

We evaluate the state-of-the-art multimodal "visual semantic" model CLIP ("Contrastive Language Image Pretraining") for biases related to the marking of age, gender, and race or ethnicity. Given the option to label an image as "a photo of a person" or to select a label denoting race or ethnicity, CLIP chooses the "person" label 47.9% of the time for White individuals, compared with 5.0% or less for individuals who are Black, East Asian, Southeast Asian, Indian, or Latino or Hispanic. The model is more likely to rank the unmarked "person" label higher than labels denoting gender for Male individuals (26.7% of the time) vs. Female individuals (15.2% of the time). Age affects whether an individual is marked by the model: Female individuals under the age of 20 are more likely than Male individuals to be marked with a gender label, but less likely to be marked with an age label, while Female individuals over the age of 40 are more likely to be marked based on age than Male individuals. We also examine the self-similarity (mean pairwise cosine similarity) for each social group, where higher self-similarity denotes greater attention directed by CLIP to the shared characteristics (age, race, or gender) of the social group. As age increases, the self-similarity of representations of Female individuals increases at a higher rate than for Male individuals, with the disparity most pronounced at the "more than 70" age range. All ten of the most self-similar social groups are individuals under the age of 10 or over the age of 70, and six of the ten are Female individuals. Existing biases of self-similarity and markedness between Male and Female gender groups are further exacerbated when the groups compared are individuals who are White and Male and individuals who are Black and Female. Results indicate that CLIP reflects the biases of the language and society which produced its training data.

📄 PDF Abstract BibTeX arXiv:2205.11378

Code (1)

wolferobert3/visual_semantic_markedness 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Modeling Markedness with a Split-and-Merger Model of Sound Change

2019-08-01 · WS 2019 8 · Andrea Ceolin, Ollie Sayeed

The concept of {`}markedness{'} has been influential in phonology for almost a century. Theoretical phonology has found it useful to describe some segments as more {`}marked{'} than others, referring to a cluster of lang…

The Overall Markedness of Discourse Relations

2015-09-01 · EMNLP 2015 9 · Lifeng Jin, Marie-Catherine de Marneffe

Evaluation: from precision, recall and F-measure to ROC, informedness, markedness and correlation

2020-10-11 · David M. W. Powers

Commonly used evaluation measures including Recall, Precision, F-Measure and Rand Accuracy are biased and should not be used without clear understanding of the biases, and corresponding identification of chance or base c…

Visualizing and Understanding Neural Models in NLP

2015-06-02 · NAACL 2016 6 · Jiwei Li, Xinlei Chen, Eduard Hovy, Dan Jurafsky

While neural networks have been successfully applied to many NLP tasks the resulting vector-based models are very difficult to interpret. For example it's not clear how they achieve {\em compositionality}, building sente…

NegationSentence

Accommodation of Conversational Code-Choice

2018-07-01 · WS 2018 7 · Anshul Bawa, Monojit Choudhury, Kalika Bali

Bilingual speakers often freely mix languages. However, in such bilingual conversations, are the language choices of the speakers coordinated? How much does one speaker{'}s choice of language affect other speakers? In th…

Information RetrievalRetrieval