paper-with-me

홈 › Papers

Investigating Use Cases of AI-Powered Scene Description Applications for Blind and Low Vision People

2024-03-22 · Ricardo Gonzalez, Jazmin Collins, Shiri Azenkot, Cynthia Bennett

"Scene description" applications that describe visual content in a photo are useful daily tools for blind and low vision (BLV) people. Researchers have studied their use, but they have only explored those that leverage remote sighted assistants; little is known about applications that use AI to generate their descriptions. Thus, to investigate their use cases, we conducted a two-week diary study where 16 BLV participants used an AI-powered scene description application we designed. Through their diary entries and follow-up interviews, users shared their information goals and assessments of the visual descriptions they received. We analyzed the entries and found frequent use cases, such as identifying visual features of known objects, and surprising ones, such as avoiding contact with dangerous objects. We also found users scored the descriptions relatively low on average, 2.76 out of 5 (SD=1.49) for satisfaction and 2.43 out of 4 (SD=1.16) for trust, showing that descriptions still need significant improvements to deliver satisfying and trustworthy experiences. We discuss future opportunities for AI as it becomes a more powerful accessibility tool for BLV users.

📄 PDF Abstract BibTeX arXiv:2403.15604

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Seeing things or seeing scenes: Investigating the capabilities of V&L models to align scene descriptions to images

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Images can be described in terms of the objects they contain, or in terms of the types of scene or place that they instantiate. In this paper we address to what extent pretrained Vision and Language models can learn to a…

Object

Sketchforme: Composing Sketched Scenes from Text Descriptions for Interactive Applications

2019-04-08 · Forrest Huang, John F. Canny

Sketching and natural languages are effective communication media for interactive applications. We introduce Sketchforme, the first neural-network-based system that can generate sketches based on text descriptions specif…

Text-to-Scene with Large Reasoning Models

2025-09-30 · Frédéric Berdoz, Luca A. Lanzendörfer, Nick Tuninga, Roger Wattenhofer arxiv

Prompt-driven scene synthesis allows users to generate complete 3D environments from textual descriptions. Current text-to-scene methods often struggle with complex geometries and object transformations, and tend to show…

Spatial ReasoningScene Generation

AIris: An AI-powered Wearable Assistive Device for the Visually Impaired

2024-05-13 · Dionysia Danai Brilli, Evangelos Georgaras, Stefania Tsilivaki, Nikos Melanitis 외

Assistive technologies for the visually impaired have evolved to facilitate interaction with a complex and dynamic world. In this paper, we introduce AIris, an AI-powered wearable device that provides environmental aware…

Face RecognitionObject Recognition

NightAdapter: Learning a Frequency Adapter for Generalizable Night-time Scene Segmentation

2025-01-01 · CVPR 2025 1 · Qi Bi, Jingjun Yi, Huimin Huang, Hao Zheng 외

Night-time scene segmentation is a critical yet challenging task in the real-world applications, primarily due to the complicated lighting conditions. However, existing methods lack sufficient generalization ability …

Scene Segmentation