paper-with-me

홈 › Papers

Captioning Images Taken by People Who Are Blind

2020-02-20 · ECCV 2020 8 · Danna Gurari, Yinan Zhao, Meng Zhang, Nilavra Bhattacharya

While an important problem in the vision community is to design algorithms that can automatically caption images, few publicly-available datasets for algorithm development directly address the interests of real users. Observing that people who are blind have relied on (human-based) image captioning services to learn about images they take for nearly a decade, we introduce the first image captioning dataset to represent this real use case. This new dataset, which we call VizWiz-Captions, consists of over 39,000 images originating from people who are blind that are each paired with five captions. We analyze this dataset to (1) characterize the typical captions, (2) characterize the diversity of content found in the images, and (3) compare its content to that found in eight popular vision datasets. We also analyze modern image captioning algorithms to identify what makes this new dataset challenging for the vision community. We publicly-share the dataset with captioning challenge instructions at https://vizwiz.org

📄 PDF Abstract BibTeX arXiv:2002.08565

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityImage Captioning

Similar Papers 제목 키워드 기반

Quality-agnostic Image Captioning to Safely Assist People with Vision Impairment

2023-04-28 · Lu Yu, Malvina Nikandrou, Jiali Jin, Verena Rieser

Automated image captioning has the potential to be a useful tool for people with vision impairments. Images taken by this user group are often noisy, which leads to incorrect and even unsafe model predictions. In this pa…

Data AugmentationImage Captioning

Multi-Modal Image Captioning for the Visually Impaired

2021-05-17 · NAACL 2021 4 · Hiba Ahsan, Nikita Bhalla, Daivat Bhatt, Kaivankumar Shah

One of the ways blind people understand their surroundings is by clicking images and relying on descriptions generated by image captioning systems. Current work on captioning images for the visually impaired do not use t…

Image Captioning

VizWiz-Priv: A Dataset for Recognizing the Presence and Purpose of Private Visual Information in Images Taken by Blind People

2019-06-01 · CVPR 2019 6 · Danna Gurari, Qing Li, Chi Lin, Yinan Zhao 외

We introduce the first visual privacy dataset originating from people who are blind in order to better understand their privacy disclosures and to encourage the development of algorithms that can assist in preventing the…

Towards Real Time Egocentric Segment Captioning for The Blind and Visually Impaired in RGB-D Theatre Images

2023-08-26 · Khadidja Delloul, Slimane Larabi

In recent years, image captioning and segmentation have emerged as crucial tasks in computer vision, with applications ranging from autonomous driving to content analysis. Although multiple solutions have emerged to help…

Autonomous DrivingImage Captioning

Assessing Image Quality Issues for Real-World Problems

2020-03-27 · CVPR 2020 6 · Tai-Yin Chiu, Yinan Zhao, Danna Gurari

We introduce a new large-scale dataset that links the assessment of image quality issues to two practical vision tasks: image captioning and visual question answering. First, we identify for 39,181 images taken by people…

Image CaptioningQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)