paper-with-me

홈 › Papers

Text-to-Image Generation for Abstract Concepts

2023-09-26 · Jiayi Liao, Xu Chen, Qiang Fu, Lun Du, Xiangnan He, Xiang Wang, Shi Han, Dongmei Zhang

Recent years have witnessed the substantial progress of large-scale models across various domains, such as natural language processing and computer vision, facilitating the expression of concrete concepts. Unlike concrete concepts that are usually directly associated with physical objects, expressing abstract concepts through natural language requires considerable effort, which results from their intricate semantics and connotations. An alternative approach is to leverage images to convey rich visual information as a supplement. Nevertheless, existing Text-to-Image (T2I) models are primarily trained on concrete physical objects and tend to fail to visualize abstract concepts. Inspired by the three-layer artwork theory that identifies critical factors, intent, object and form during artistic creation, we propose a framework of Text-to-Image generation for Abstract Concepts (TIAC). The abstract concept is clarified into a clear intent with a detailed definition to avoid ambiguity. LLMs then transform it into semantic-related physical objects, and the concept-dependent form is retrieved from an LLM-extracted form pattern set. Information from these three aspects will be integrated to generate prompts for T2I models via LLM. Evaluation results from human assessments and our newly designed metric concept score demonstrate the effectiveness of our framework in creating images that can sufficiently express abstract concepts.

📄 PDF Abstract BibTeX arXiv:2309.14623

Code (0)

등록된 구현이 없습니다.

Tasks

FormImage GenerationText to Image GenerationText-to-Image Generation

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

Weakly Supervised Annotations for Multi-modal Greeting Cards Dataset

2022-12-01 · Sidra Hanif, Longin Jan Latecki

In recent years, there is a growing number of pre-trained models trained on a large corpus of data and yielding good performance on various tasks such as classifying multimodal datasets. These models have shown good perf…

Image CaptioningImage GenerationText to Image GenerationText-to-Image Generation

Mod-Adapter: Tuning-Free and Versatile Multi-concept Personalization via Modulation Adapter

2025-05-24 · Weizhi Zhong, Huan Yang, Zheng Liu, Huiguo He 외

Personalized text-to-image generation aims to synthesize images of user-provided concepts in diverse contexts. Despite recent progress in multi-concept personalization, most are limited to object concepts and struggle to…

Image GenerationMixture-of-ExpertsText to Image GenerationText-to-Image Generation

Bongard-RWR+: Real-World Representations of Fine-Grained Concepts in Bongard Problems

2025-08-16 · Szymon Pawlonka, Mikołaj Małkiński, Jacek Mańdziuk arxiv

Bongard Problems (BPs) provide a challenging testbed for abstract visual reasoning (AVR), requiring models to identify visual concepts fromjust a few examples and describe them in natural language. Early BP benchmarks fe…

Answer GenerationVisual Reasoning

META4: Semantically-Aligned Generation of Metaphoric Gestures Using Self-Supervised Text and Speech Representation

2023-11-09 · Mireille Fares, Catherine Pelachaud, Nicolas Obin

Image Schemas are repetitive cognitive patterns that influence the way we conceptualize and reason about various concepts present in speech. These patterns are deeply embedded within our cognitive processes and are refle…

Describing Natural Images Containing Novel Objects with Knowledge Guided Assitance

2017-10-17 · Aditya Mogadala, Umanga Bista, Lexing Xie, Achim Rettinger

Images in the wild encapsulate rich knowledge about varied abstract concepts and cannot be sufficiently described with models built only using image-caption pairs containing selected objects. We propose to handle such a …

Caption Generation