paper-with-me

홈 › Papers

Leveraging Text-Driven Semantic Variation for Robust OOD Segmentation

2025-11-10 · Seungheon Song, Jaekoo Lee arxiv

In autonomous driving and robotics, ensuring road safety and reliable decision-making critically depends on out-of-distribution (OOD) segmentation. While numerous methods have been proposed to detect anomalous objects on the road, leveraging the vision-language space-which provides rich linguistic knowledge-remains an underexplored field. We hypothesize that incorporating these linguistic cues can be especially beneficial in the complex contexts found in real-world autonomous driving scenarios. To this end, we present a novel approach that trains a Text-Driven OOD Segmentation model to learn a semantically diverse set of objects in the vision-language space. Concretely, our approach combines a vision-language model's encoder with a transformer decoder, employs Distance-Based OOD prompts located at varying semantic distances from in-distribution (ID) classes, and utilizes OOD Semantic Augmentation for OOD representations. By aligning visual and textual information, our approach effectively generalizes to unseen objects and provides robust OOD segmentation in diverse driving environments. We conduct extensive experiments on publicly available OOD segmentation datasets such as Fishyscapes, Segment-Me-If-You-Can, and Road Anomaly datasets, demonstrating that our approach achieves state-of-the-art performance across both pixel-level and object-level evaluations. This result underscores the potential of vision-language-based OOD segmentation to bolster the safety and reliability of future autonomous driving systems.

📄 PDF Abstract BibTeX arXiv:2511.07238

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

Multi-Text Guided Few-Shot Semantic Segmentation

2025-11-19 · Qiang Jiao, Bin Yan, Yi Yang, Mengrui Shi 외 arxiv

Recent CLIP-based few-shot semantic segmentation methods introduce class-level textual priors to assist segmentation by typically using a single prompt (e.g., a photo of class). However, these approaches often result in …

Few-Shot Semantic Segmentation

3D-STMN: Dependency-Driven Superpoint-Text Matching Network for End-to-End 3D Referring Expression Segmentation

2023-08-31 · Changli Wu, Yiwei Ma, Qi Chen, Haowei Wang 외

In 3D Referring Expression Segmentation (3D-RES), the earlier approach adopts a two-stage paradigm, extracting segmentation proposals and then matching them with referring expressions. However, this conventional paradigm…

NavigateReferring ExpressionReferring Expression SegmentationSegmentation+1

Can Textual Semantics Mitigate Sounding Object Segmentation Preference?

2024-07-15 · Yaoting Wang, Peiwen Sun, Yuanchao Li, Honggang Zhang 외

The Audio-Visual Segmentation (AVS) task aims to segment sounding objects in the visual space using audio cues. However, in this work, it is recognized that previous AVS methods show a heavy reliance on detrimental segme…

Language ModellingLarge Language ModelObjectSegmentation+1

Semantic Browsing: Controllable Diversity for Image Generation

2026-06-22 · Sara Dorfman, Maya Vishnevsky, Omer Dahary, Or Patashnik 외 arxiv

Modern text-to-image models excel in visual fidelity and prompt adherence. However, this strict adherence comes at the cost of diversity: generated samples tend to collapse into a single visual interpretation. Existing m…

Image Generation

Deep learning and its application to medical image segmentation

2018-03-23 · Holger R. Roth, Chen Shen, Hirohisa ODA, Masahiro Oda 외

One of the most common tasks in medical imaging is semantic segmentation. Achieving this segmentation automatically has been an active area of research, but the task has been proven very challenging due to the large vari…

AnatomyComputed Tomography (CT)Deep LearningImage Segmentation+4