paper-with-me

Papers

Generating Contextually-Relevant Navigation Instructions for Blind and Low Vision People

2024-07-11 · Zain Merchant, Abrar Anwar, Emily Wang, Souti Chattopadhyay, Jesse Thomason

Navigating unfamiliar environments presents significant challenges for blind and low-vision (BLV) individuals. In this work, we construct a dataset of images and goals across different scenarios such as searching through kitchens or navigating outdoors. We then investigate how grounded instruction generation methods can provide contextually-relevant navigational guidance to users in these instances. Through a sighted user study, we demonstrate that large pretrained language models can produce correct and useful instructions perceived as beneficial for BLV users. We also conduct a survey and interview with 4 BLV users and observe useful insights on preferences for different instructions based on the scenario.

📄 PDF Abstract BibTeX arXiv:2407.08219

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SeqWalker: Sequential-Horizon Vision-and-Language Navigation with Hierarchical Planning

2026-01-08 · Zebin Han, Xudong Wang, Baichen Liu, Qi Lyu 외 arxiv

Sequential-Horizon Vision-and-Language Navigation (SH-VLN) presents a challenging scenario where agents should sequentially execute multi-task navigation guided by complex, long-horizon language instructions. Current vis…

iMotion-LLM: Motion Prediction Instruction Tuning

2024-06-10 · Abdulwahab Felemban, Eslam Mohamed BAKR, Xiaoqian Shen, Jian Ding 외

We introduce iMotion-LLM: a Multimodal Large Language Models (LLMs) with trajectory prediction, tailored to guide interactive multi-agent scenarios. Different from conventional motion prediction approaches, iMotion-LLM c…

Autonomous Navigationmotion predictionPredictionTrajectory Prediction

An Efficient Indoor Navigation Technique To Find Optimal Route For Blinds Using QR Codes

2020-05-29 · Affan Idrees, Zahid Iqbal, Maria Ishfaq

Blind navigation is an accessibility application that enables blind to use an android Smartphone in an easy way for indoor navigation with instructions in audio form. We have proposed a prototype which is an indoor navig…

NavRAG: Generating User Demand Instructions for Embodied Navigation through Retrieval-Augmented LLM

2025-02-16 · Zihan Wang, Yaohui Zhu, Gim Hee Lee, Yachun Fan

Vision-and-Language Navigation (VLN) is an essential skill for embodied agents, allowing them to navigate in 3D environments following natural language instructions. High-performance navigation models require a large amo…

NavigateRAGRetrievalRetrieval-augmented Generation+3

Generating Landmark Navigation Instructions from Maps as a Graph-to-Text Problem

2020-12-30 · ACL 2021 5 · Raphael Schumann, Stefan Riezler

Car-focused navigation services are based on turns and distances of named streets, whereas navigation instructions naturally used by humans are centered around physical objects called landmarks. We present a neural model…

Natural Language Landmark Navigation Instructions Generation