paper-with-me

Papers

Identifying Crucial Objects in Blind and Low-Vision Individuals' Navigation

2024-08-23 · Md Touhidul Islam, Imran Kabir, Elena Ariel Pearce, Md Alimoor Reza, Syed Masum Billah

This paper presents a curated list of 90 objects essential for the navigation of blind and low-vision (BLV) individuals, encompassing road, sidewalk, and indoor environments. We develop the initial list by analyzing 21 publicly available videos featuring BLV individuals navigating various settings. Then, we refine the list through feedback from a focus group study involving blind, low-vision, and sighted companions of BLV individuals. A subsequent analysis reveals that most contemporary datasets used to train recent computer vision models contain only a small subset of the objects in our proposed list. Furthermore, we provide detailed object labeling for these 90 objects across 31 video segments derived from the original 21 videos. Finally, we make the object list, the 21 videos, and object labeling in the 31 video segments publicly available. This paper aims to fill the existing gap and foster the development of more inclusive and effective navigation aids for the BLV community.

📄 PDF Abstract BibTeX arXiv:2408.13175

Code (0)

등록된 구현이 없습니다.

Tasks

Object

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation

2024-07-23 · Md Touhidul Islam, Imran Kabir, Elena Ariel Pearce, Md Alimoor Reza 외

This paper introduces a dataset for improving real-time object recognition systems to aid blind and low-vision (BLV) individuals in navigation tasks. The dataset comprises 21 videos of BLV individuals navigating outdoor …

Object Recognition

Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models

2024-05-28 · Sangmin Woo, Donguk Kim, Jaehyuk Jang, Yubin Choi 외

This study addresses the issue observed in Large Vision Language Models (LVLMs), where excessive attention on a few image tokens, referred to as blind tokens, leads to hallucinatory responses in tasks requiring fine-grai…

MMEObject

Enhancing Community Vision Screening -- AI Driven Retinal Photography for Early Disease Detection and Patient Trust

2024-10-27 · Xiaofeng Lei, Yih-Chung Tham, Jocelyn Hui Lin Goh, Yangqin Feng 외

Community vision screening plays a crucial role in identifying individuals with vision loss and preventing avoidable blindness, particularly in rural communities where access to eye care services is limited. Currently, t…

Small Object Detection for Indoor Assistance to the Blind using YOLO NAS Small and Super Gradients

2024-08-28 · Rashmi BN, R. Guru, Anusuya M A

Advancements in object detection algorithms have opened new avenues for assistive technologies that cater to the needs of visually impaired individuals. This paper presents a novel approach for indoor assistance to the b…

Objectobject-detectionObject DetectionSmall Object Detection

A Light and Smart Wearable Platform with Multimodal Foundation Model for Enhanced Spatial Reasoning in People with Blindness and Low Vision

2025-05-16 · Alexey Magay, Dhurba Tripathi, Yu Hao, Yi Fang

People with blindness and low vision (pBLV) face significant challenges, struggling to navigate environments and locate objects due to limited visual cues. Spatial reasoning is crucial for these individuals, as it enable…

Large Language ModelNavigateObject RecognitionSpatial Reasoning