paper-with-me

홈 › Papers

Generalized People Diversity: Learning a Human Perception-Aligned Diversity Representation for People Images

2024-01-25 · Hansa Srinivasan, Candice Schumann, Aradhana Sinha, David Madras, Gbolahan Oluwafemi Olanubi, Alex Beutel, Susanna Ricco, Jilin Chen

Capturing the diversity of people in images is challenging: recent literature tends to focus on diversifying one or two attributes, requiring expensive attribute labels or building classifiers. We introduce a diverse people image ranking method which more flexibly aligns with human notions of people diversity in a less prescriptive, label-free manner. The Perception-Aligned Text-derived Human representation Space (PATHS) aims to capture all or many relevant features of people-related diversity, and, when used as the representation space in the standard Maximal Marginal Relevance (MMR) ranking algorithm, is better able to surface a range of types of people-related diversity (e.g. disability, cultural attire). PATHS is created in two stages. First, a text-guided approach is used to extract a person-diversity representation from a pre-trained image-text model. Then this representation is fine-tuned on perception judgments from human annotators so that it captures the aspects of people-related similarity that humans find most salient. Empirical results show that the PATHS method achieves diversity better than baseline methods, according to side-by-side ratings from human annotators.

📄 PDF Abstract BibTeX arXiv:2401.14322

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeDiversity

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

AlignFace: Human-Aligned Face Similarity Metric with Interpretable Concept Relations

2026-08-14 · Ying Huang, Wencan Zhang, Brian Y. Lim arxiv

Computer vision models for generated facial content, such as face editing and privacy protection, increasingly affect people, requiring similarity metrics that serve as faithful proxies for human perception. While percep…

Personalized Image Descriptions from Attention Sequences

2025-12-07 · Ruoyu Xue, Hieu Le, Jingyi Xu, Sounak Mondal 외 arxiv

People can view the same image differently: they focus on different regions, objects, and details in varying orders and describe them in distinct linguistic styles. This leads to substantial variability in image descript…

Dimensions of Diversity in Human Perceptions of Algorithmic Fairness

2020-05-02 · Nina Grgić-Hlača, Gabriel Lima, Adrian Weller, Elissa M. Redmiles

A growing number of oversight boards and regulatory bodies seek to monitor and govern algorithms that make decisions about people's lives. Prior work has explored how people believe algorithmic decisions should be made, …

Decision MakingDiversityFairness

Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function

2024-06-03 · Keyon Vafa, Ashesh Rambachan, Sendhil Mullainathan

What makes large language models (LLMs) impressive is also what makes them hard to evaluate: their diversity of uses. To evaluate these models, we must understand the purposes they will be used for. We consider a setting…

DiversityMMLU

AI and My Values: User Perceptions of LLMs' Ability to Extract, Embody, and Explain Human Values from Casual Conversations

2026-01-30 · Bhada Yun, Renn Su, April Yi Wang arxiv

Does AI understand human values? While this remains an open philosophical question, we take a pragmatic stance by introducing VAPT, the Value-Alignment Perception Toolkit, for studying how LLMs reflect people's values an…