paper-with-me

홈 › Papers

Exploring Language Patterns of Prompts in Text-to-Image Generation and Their Impact on Visual Diversity

2025-04-19 · Maria-Teresa De Rosa Palmini, Eva Cetinic

Following the initial excitement, Text-to-Image (TTI) models are now being examined more critically. While much of the discourse has focused on biases and stereotypes embedded in large-scale training datasets, the sociotechnical dynamics of user interactions with these models remain underexplored. This study examines the linguistic and semantic choices users make when crafting prompts and how these choices influence the diversity of generated outputs. Analyzing over six million prompts from the Civiverse dataset on the CivitAI platform across seven months, we categorize users into three groups based on their levels of linguistic experimentation: consistent repeaters, occasional repeaters, and non-repeaters. Our findings reveal that as user participation grows over time, prompt language becomes increasingly homogenized through the adoption of popular community tags and descriptors, with repeated prompts comprising 40-50% of submissions. At the same time, semantic similarity and topic preferences remain relatively stable, emphasizing common subjects and surface aesthetics. Using Vendi scores to quantify visual diversity, we demonstrate a clear correlation between lexical similarity in prompts and the visual similarity of generated images, showing that linguistic repetition reinforces less diverse representations. These findings highlight the significant role of user-driven factors in shaping AI-generated imagery, beyond inherent model biases, and underscore the need for tools and practices that encourage greater linguistic and thematic experimentation within TTI systems to foster more inclusive and diverse AI-generated content.

📄 PDF Abstract BibTeX arXiv:2504.14125

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityImage GenerationSemantic SimilaritySemantic Textual SimilarityText to Image GenerationText-to-Image Generation

Similar Papers 제목 키워드 기반

Exploring Prompt-Based Methods for Zero-Shot Hypernym Prediction with Large Language Models

2024-01-09 · Mikhail Tikhomirov, Natalia Loukachevitch

This article investigates a zero-shot approach to hypernymy prediction using large language models (LLMs). The study employs a method based on text probability calculation, applying it to various generated prompts. The e…

Language ModelingLanguage Modelling

MCP2OSC: Parametric Control by Natural Language

2025-08-14 · Yuan-Yi Fan arxiv

Text prompts enable intuitive content creation but may fall short in achieving high precision for intricate tasks; knob or slider controls offer precise adjustments at the cost of increased complexity. To address the gap…

IteraTTA: An interface for exploring both text prompts and audio priors in generating music with text-to-audio models

2023-07-24 · Hiromu Yakura, Masataka Goto

Recent text-to-audio generation techniques have the potential to allow novice users to freely generate music audio. Even if they do not have musical knowledge, such as about chord progressions and instruments, users can …

Audio GenerationMusic Generation

Exploring the Curious Case of Code Prompts

2023-04-26 · Li Zhang, Liam Dugan, Hainiu Xu, Chris Callison-Burch

Recent work has shown that prompting language models with code-like representations of natural language leads to performance improvements on structured reasoning tasks. However, such tasks comprise only a small subset of…

Generative AI in Color-Changing Systems: Re-Programmable 3D Object Textures with Material and Design Constraints

2024-04-25 · Yunyi Zhu, Faraz Faruqi, Stefanie Mueller

Advances in Generative AI tools have allowed designers to manipulate existing 3D models using text or image-based prompts, enabling creators to explore different design goals. Photochromic color-changing systems, on the …