paper-with-me

홈 › Papers

ColorConceptBench: A Benchmark for Probabilistic Color-Concept Understanding in Text-to-Image Models

2026-01-23 · Chenxi Ruan, Yihan Hou, Yu Xiao, Guosheng Hu, Wei Zeng arxiv

Text-to-image (T2I) models have advanced considerably in generating high-quality images from textual descriptions. However, their ability to associate colors with concepts remains largely constrained to explicit color names or codes, while their capacity to handle \emph{implicit concepts}, such as emotions and visual states, remains underexplored. To address this gap, we introduce ColorConceptBench, an expert-annotated benchmark that systematically evaluates color-concept associations through probabilistic color distributions. ColorConceptBench moves beyond explicit color specifications by examining how models interpret 1,281 implicit color concepts, grounded in 6,584 human annotations. Our evaluation of nine leading T2I models reveals that performance varies substantially across semantic categories, and models exhibit a significant lack of sensitivity to abstract semantics. These limitations persist even when applying classifier-free guidance scaling at inference time, suggesting that achieving human-like color understanding demands a shift in how models learn and represent implicit semantic meaning.

📄 PDF Abstract BibTeX arXiv:2601.16836

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Probing Conceptual Understanding of Large Visual-Language Models

2023-04-07 · Madeline Schiappa, Raiyaan Abdullah, Shehreen Azad, Jared Claypoole 외

In recent years large visual-language (V+L) models have achieved great success in various downstream tasks. However, it is not well studied whether these models have a conceptual grasp of the visual content. In this work…

Benchmarking

Shades of confusion: Lexical uncertainty modulates ad hoc coordination in an interactive communication task

2021-05-13 · Sonia K. Murthy, Thomas L. Griffiths, Robert D. Hawkins

There is substantial variability in the expectations that communication partners bring into interactions, creating the potential for misunderstandings. To directly probe these gaps and our ability to overcome them, we pr…

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

2026-02-26 · Hongyu Li, Kuan Liu, Yuan Chen, Juntao Hu 외 arxiv

Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a "Paradox of Simplicity": models that can render complex scenes often fail at trivial, low-entropy ta…

Image Generation

A knowledge-based intelligent system for control of dirt recognition process in the smart washing machines

2019-05-02 · Mohsen Annabestani, Alireza Rowhanimanesh, Akram Rezaei, Ladan Avazpour 외

In this paper, we propose an intelligence approach based on fuzzy logic to modeling human intelligence in washing clothes. At first, an intelligent feedback loop is designed for perception-based sensing of dirt inspired …

Decision Making

Color in Visual-Language Models: CLIP deficiencies

2025-02-06 · Guillem Arias, Ramon Baldrich, Maria Vanrell

This work explores how color is encoded in CLIP (Contrastive Language-Image Pre-training) which is currently the most influential VML (Visual Language model) in Artificial Intelligence. After performing different experim…

Attribute