paper-with-me

홈 › Papers

GenComUI: Exploring Generative Visual Aids as Medium to Support Task-Oriented Human-Robot Communication

2025-02-15 · Yate Ge, Meiying Li, Xipeng Huang, Yuanda Hu, Qi Wang, Xiaohua Sun, Weiwei Guo

This work investigates the integration of generative visual aids in human-robot task communication. We developed GenComUI, a system powered by large language models that dynamically generates contextual visual aids (such as map annotations, path indicators, and animations) to support verbal task communication and facilitate the generation of customized task programs for the robot. This system was informed by a formative study that examined how humans use external visual tools to assist verbal communication in spatial tasks. To evaluate its effectiveness, we conducted a user experiment (n = 20) comparing GenComUI with a voice-only baseline. The results demonstrate that generative visual aids, through both qualitative and quantitative analysis, enhance verbal task communication by providing continuous visual feedback, thus promoting natural and effective human-robot communication. Additionally, the study offers a set of design implications, emphasizing how dynamically generated visual aids can serve as an effective communication medium in human-robot interaction. These findings underscore the potential of generative visual aids to inform the design of more intuitive and effective human-robot communication, particularly for complex communication scenarios in human-robot interaction and LLM-based end-user development.

📄 PDF Abstract BibTeX arXiv:2502.10678

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts

2025-08-30 · Adam Cole, Mick Grierson arxiv

This paper presents an artistic and technical investigation into the attention mechanisms of video diffusion transformers. Inspired by early video artists who manipulated analog video signals to create new visual aesthet…

Text-to-Video Generation

Telling Creative Stories Using Generative Visual Aids

2021-10-27 · Safinah Ali, Devi Parikh

Can visual artworks created using generative visual algorithms inspire human creativity in storytelling? We asked writers to write creative stories from a starting prompt, and provided them with visuals created by genera…

VISTA: Visual Integrated System for Tailored Automation in Math Problem Generation Using LLM

2024-11-08 · Jeongwoo Lee, Kwangsuk Park, Jihyeon Park

Generating accurate and consistent visual aids is a critical challenge in mathematics education, where visual representations like geometric shapes and functions play a pivotal role in enhancing student comprehension. Th…

Math

Art and the science of generative AI: A deeper dive

2023-06-07 · Ziv Epstein, Aaron Hertzmann, Laura Herman, Robert Mahari 외

A new class of tools, colloquially called generative AI, can produce high-quality artistic media for visual arts, concept art, music, fiction, literature, video, and animation. The generative capabilities of these tools …

Thinking with Constructions: A Benchmark and Policy Optimization for Visual-Text Interleaved Geometric Reasoning

2026-03-19 · Haokun Zhao, Wanshi Xu, Haidong Yuan, Songjun Cao 외 arxiv

Geometric reasoning inherently requires "thinking with constructions" -- the dynamic manipulation of visual aids to bridge the gap between problem conditions and solutions. However, existing Multimodal Large Language Mod…

Reinforcement Learning