paper-with-me

홈 › Papers

CoT-Segmenter: Enhancing OOD Detection in Dense Road Scenes via Chain-of-Thought Reasoning

2025-07-05 · Jeonghyo Song, Kimin Yun, DaeUng Jo, Jinyoung Kim, Youngjoon Yoo arxiv

Effective Out-of-Distribution (OOD) detection is criti-cal for ensuring the reliability of semantic segmentation models, particularly in complex road environments where safety and accuracy are paramount. Despite recent advancements in large language models (LLMs), notably GPT-4, which significantly enhanced multimodal reasoning through Chain-of-Thought (CoT) prompting, the application of CoT-based visual reasoning for OOD semantic segmentation remains largely unexplored. In this paper, through extensive analyses of the road scene anomalies, we identify three challenging scenarios where current state-of-the-art OOD segmentation methods consistently struggle: (1) densely packed and overlapping objects, (2) distant scenes with small objects, and (3) large foreground-dominant objects. To address the presented challenges, we propose a novel CoT-based framework targeting OOD detection in road anomaly scenes. Our method leverages the extensive knowledge and reasoning capabilities of foundation models, such as GPT-4, to enhance OOD detection through improved image understanding and prompt-based reasoning aligned with observed problematic scene attributes. Extensive experiments show that our framework consistently outperforms state-of-the-art methods on both standard benchmarks and our newly defined challenging subset of the RoadAnomaly dataset, offering a robust and interpretable solution for OOD semantic segmentation in complex driving environments.

📄 PDF Abstract BibTeX arXiv:2507.03984

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic SegmentationMultimodal ReasoningVisual Reasoning

Similar Papers 제목 키워드 기반

Crowd-SAM: SAM as a Smart Annotator for Object Detection in Crowded Scenes

2024-07-16 · Zhi Cai, Yingjie Gao, Yaoyan Zheng, Nan Zhou 외

In computer vision, object detection is an important task that finds its application in many scenarios. However, obtaining extensive labels can be challenging, especially in crowded scenes. Recently, the Segment Anything…

Human Instance SegmentationInstance Segmentationobject-detectionObject Detection+1

A Simple Video Segmenter by Tracking Objects Along Axial Trajectories

2023-11-30 · Ju He, Qihang Yu, Inkyu Shin, Xueqing Deng 외

Video segmentation requires consistently segmenting and tracking objects over time. Due to the quadratic dependency on input size, directly applying self-attention to video segmentation with high-resolution input feature…

GPUObjectObject TrackingPanoptic Segmentation+5

No More Discrimination: Cross City Adaptation of Road Scene Segmenters

2017-04-27 · ICCV 2017 10 · Yi-Hsin Chen, Wei-Yu Chen, Yu-Ting Chen, Bo-Cheng Tsai 외

Despite the recent success of deep-learning based semantic segmentation, deploying a pre-trained road scene segmenter to a city whose images are not presented in the training set would not achieve satisfactory performanc…

SegmentationSemantic Segmentation

Concealed Object Segmentation with Hierarchical Coherence Modeling

2024-01-22 · Fengyang Xiao, Pan Zhang, Chunming He, Runze Hu 외

Concealed object segmentation (COS) is a challenging task that involves localizing and segmenting those concealed objects that are visually blended with their surrounding environments. Despite achieving remarkable succes…

DecoderImage SegmentationObjectobject-detection+4

Prediction of Occluded Pedestrians in Road Scenes using Human-like Reasoning: Insights from the OccluRoads Dataset

2024-12-09 · Melo Castillo Angie Nataly, Martin Serrano Sergio, Salinas Carlota, Sotelo Miguel Angel

Pedestrian detection is a critical task in autonomous driving, aimed at enhancing safety and reducing risks on the road. Over recent years, significant advancements have been made in improving detection performance. Howe…

Autonomous DrivingBayesian InferenceGraph EmbeddingKnowledge Graph Embedding+1