Uncertainty-Gated Region-Level Retrieval for Robust Semantic Segmentation
Semantic segmentation of outdoor street scenes plays a key role in applications such as autonomous driving, mobile robotics, and assistive technology for visually-impaired pedestrians. For these applications, accurately distinguishing between key surfaces and objects such as roads, sidewalks, vehicles, and pedestrians is essential for maintaining safety and minimizing risks. Semantic segmentation must be robust to different environments, lighting and weather conditions, and sensor noise, while being performed in real-time. We propose a region-level, uncertainty-gated retrieval mechanism that improves segmentation accuracy and calibration under domain shift. Our best method achieves an 11.3% increase in mean intersection-over-union while reducing retrieval cost by 87.5%, retrieving for only 12.5% of regions compared to 100% for always-on baseline.
Code (0)
등록된 구현이 없습니다.
Tasks
Semantic SegmentationAutonomous DrivingSimilar Papers 제목 키워드 기반
Exploiting Visual Semantic Reasoning for Video-Text Retrieval
Video retrieval is a challenging research topic bridging the vision and language areas and has attracted broad attention in recent years. Previous works have been devoted to representing videos by directly encoding from …
RetrievalText Retrievaltext similarityVideo Retrieval+1Hierarchical Matching and Reasoning for Multi-Query Image Retrieval
As a promising field, Multi-Query Image Retrieval (MQIR) aims at searching for the semantically relevant image given multiple region-specific text queries. Existing works mainly focus on a single-level similarity between…
Image RetrievalRetrievalBayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering
Trustworthy deployment of Agentic Retrieval-Augmented Generation (RAG) systems requires mechanisms for estimating when multi-stage reasoning pipelines may fail. This paper presents an uncertainty-aware Agentic Retrieval-…
Multi-hop Question AnsweringHard to See, Hard to Label: Generative and Symbolic Acquisition for Subtle Visual Phenomena
Subtle visual anomalies such as hairline cracks, sub-millimeter voids, and low-contrast inclusions are structurally atypical yet visually ambiguous, making them both difficult to annotate and easy to overlook during acti…
Object DetectionActive LearningPolarimetric Hierarchical Semantic Model and Scattering Mechanism Based PolSAR Image Classification
For polarimetric SAR (PolSAR) image classification, it is a challenge to classify the aggregated terrain types, such as the urban area, into semantic homogenous regions due to sharp bright-dark variations in intensity. T…
General Classificationimage-classificationImage Classification