paper-with-me

Papers

The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility

2025-08-11 · Xiantao Zhang arxiv

Multimodal Large Language Models (MLLMs) hold immense promise as assistive technologies for the blind and visually impaired (BVI) community. However, we identify a critical failure mode that undermines their trustworthiness in real-world applications. We introduce the Escalator Problem -- the inability of state-of-the-art models to perceive an escalator's direction of travel -- as a canonical example of a deeper limitation we term Implicit Motion Blindness. This blindness stems from the dominant frame-sampling paradigm in video understanding, which, by treating videos as discrete sequences of static images, fundamentally struggles to perceive continuous, low-signal motion. As a position paper, our contribution is not a new model but rather to: (I) formally articulate this blind spot, (II) analyze its implications for user trust, and (III) issue a call to action. We advocate for a paradigm shift from purely semantic recognition towards robust physical perception and urge the development of new, human-centered benchmarks that prioritize safety, reliability, and the genuine needs of users in dynamic environments.

📄 PDF Abstract BibTeX arXiv:2508.07989

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Elevator, Escalator or Neither? Classifying Pedestrian Conveyor State Using Inertial Navigation System

2024-05-06 · Tianlang He, Zhiqiu Xia, S. -H. Gary Chan

Knowing a pedestrian's conveyor state of "elevator," "escalator," or "neither" is fundamental in many applications such as indoor navigation and people flow management. We study, for the first time, classifying the conve…

Indoor Localization

Potential Escalator-related Injury Identification and Prevention Based on Multi-module Integrated System for Public Health

2021-03-13 · Zeyu Jiao, Huan Lei, Hengshan Zong, Yingjie Cai 외

Escalator-related injuries threaten public health with the widespread use of escalators. The existing studies tend to focus on after-the-fact statistics, reflecting on the original design and use of defects to reduce the…

object-detectionObject Detection

KODIS: A Multicultural Dispute Resolution Dialogue Corpus

2025-04-17 · James Hale, Sushrita Rakshit, Kushal Chawla, Jeanne M. Brett 외

We present KODIS, a dyadic dispute resolution corpus containing thousands of dialogues from over 75 countries. Motivated by a theoretical model of culture and conflict, participants engage in a typical customer service d…

DeepBlindness: Fast Blindness Map Estimation and Blindness Type Classification for Outdoor Scene from Single Color Image

2019-11-02 · Jiaxiong Qiu, Xinyuan Yu, Guoqiang Yang, Shuaicheng Liu

Outdoor vision robotic systems and autonomous cars suffer from many image-quality issues, particularly haze, defocus blur, and motion blur, which we will define generically as "blindness issues". These blindness issues m…

General Classification

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

2026-06-08 · Kazi Noshin, Sajib Acharjee Dip, Ranat Das Prangon, Fardin Hassan Tamim 외 arxiv

Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support toward excessive validation or escalatory alignment. Existing sycophanc…

Response Generation