paper-with-me

Papers

GANSlider: How Users Control Generative Models for Images using Multiple Sliders with and without Feedforward Information

2022-02-02 · Hai Dang, Lukas Mecke, Daniel Buschek

We investigate how multiple sliders with and without feedforward visualizations influence users' control of generative models. In an online study (N=138), we collected a dataset of people interacting with a generative adversarial network (StyleGAN2) in an image reconstruction task. We found that more control dimensions (sliders) significantly increase task difficulty and user actions. Visual feedforward partly mitigates this by enabling more goal-directed interaction. However, we found no evidence of faster or more accurate task performance. This indicates a tradeoff between feedforward detail and implied cognitive costs, such as attention. Moreover, we found that visualizations alone are not always sufficient for users to understand individual control dimensions. Our study quantifies fundamental UI design factors and resulting interaction behavior in this context, revealing opportunities for improvement in the UI design for interactive applications of generative models. We close by discussing design directions and further aspects.

📄 PDF Abstract BibTeX arXiv:2202.00965

Code (0)

등록된 구현이 없습니다.

Tasks

Generative Adversarial NetworkImage Reconstruction

Similar Papers 제목 키워드 기반

SceneCraft: Interactive System for Image Editing via Scene Graph

2026-06-15 · Duc-Manh Phan, Ngoc-Dai Tran, Duy-Khang Do, Tam V. Nguyen 외 arxiv

Recent advances in generative AI have enabled natural language-driven image editing, yet existing systems often fail in complex scenes with multiple interacting objects because they rely heavily on users crafting precise…

Prompt EngineeringImage Editing

MCGM: Mask Conditional Text-to-Image Generative Model

2024-10-01 · Rami Skaik, Leonardo Rossi, Tomaso Fontanini, Andrea Prati

Recent advancements in generative models have revolutionized the field of artificial intelligence, enabling the creation of highly-realistic and detailed images. In this study, we propose a novel Mask Conditional Text-to…

model

Controlling Rate, Distortion, and Realism: Towards a Single Comprehensive Neural Image Compression Model

2024-05-27 · Shoma Iwai, Tomo Miyazaki, Shinichiro Omachi

In recent years, neural network-driven image compression (NIC) has gained significant attention. Some works adopt deep generative models such as GANs and diffusion models to enhance perceptual quality (realism). A critic…

Image Compression

InstructMesh: Selective Refinement of Generative 3D Models for Fabrication

2026-08-28 · Faraz Faruqi, Ahmed Katary, Demircan Tas, Theresa Hradilak 외 arxiv

Recent advances in generative AI allow users to create 3D models from text or images. However, these models prioritize visual plausibility over geometric accuracy, often generating results with flaws that compromise thei…

AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People

2024-08-05 · Seonghee Lee, Maho Kohga, Steve Landau, Sile O'Modhrain 외

People with visual impairments often struggle to create content that relies heavily on visual elements, particularly when conveying spatial and structural information. Existing accessible drawing tools, which construct i…

Math