paper-with-me

홈 › Papers

Panoramic Out-of-Distribution Segmentation

2025-05-06 · Mengfei Duan, Kailun Yang, Yuheng Zhang, Yihong Cao, Fei Teng, Kai Luo, Jiaming Zhang, Zhiyong Li, Shutao Li

Panoramic imaging enables capturing 360{\deg} images with an ultra-wide Field-of-View (FoV) for dense omnidirectional perception. However, current panoramic semantic segmentation methods fail to identify outliers, and pinhole Out-of-distribution Segmentation (OoS) models perform unsatisfactorily in the panoramic domain due to background clutter and pixel distortions. To address these issues, we introduce a new task, Panoramic Out-of-distribution Segmentation (PanOoS), achieving OoS for panoramas. Furthermore, we propose the first solution, POS, which adapts to the characteristics of panoramic images through text-guided prompt distribution learning. Specifically, POS integrates a disentanglement strategy designed to materialize the cross-domain generalization capability of CLIP. The proposed Prompt-based Restoration Attention (PRA) optimizes semantic decoding by prompt guidance and self-adaptive correction, while Bilevel Prompt Distribution Learning (BPDL) refines the manifold of per-pixel mask embeddings via semantic prototype supervision. Besides, to compensate for the scarcity of PanOoS datasets, we establish two benchmarks: DenseOoS, which features diverse outliers in complex environments, and QuadOoS, captured by a quadruped robot with a panoramic annular lens system. Extensive experiments demonstrate superior performance of POS, with AuPRC improving by 34.25% and FPR95 decreasing by 21.42% on DenseOoS, outperforming state-of-the-art pinhole-OoS methods. Moreover, POS achieves leading closed-set segmentation capabilities. Code and datasets will be available at https://github.com/MengfeiD/PanOoS.

📄 PDF Abstract BibTeX arXiv:2505.03539

Code (1)

mengfeid/panoos 공식 구현

Tasks

DisentanglementDomain GeneralizationPOSSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Transfer beyond the Field of View: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation

2021-10-21 · Jiaming Zhang, Chaoxiang Ma, Kailun Yang, Alina Roitberg 외

Autonomous vehicles clearly benefit from the expanded Field of View (FoV) of 360-degree sensors, but modern semantic segmentation approaches rely heavily on annotated training data which is rarely available for panoramic…

Autonomous VehiclesDomain AdaptationSegmentationSemantic Segmentation+1

Panoramic Panoptic Segmentation: Towards Complete Surrounding Understanding via Unsupervised Contrastive Learning

2021-03-01 · Alexander Jaus, Kailun Yang, Rainer Stiefelhagen

In this work, we introduce panoramic panoptic segmentation as the most holistic scene understanding both in terms of field of view and image level understanding for standard camera based input. A complete surrounding und…

Contrastive LearningPanoptic SegmentationScene UnderstandingSegmentation

DensePASS: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation with Attention-Augmented Context Exchange

2021-08-13 · Chaoxiang Ma, Jiaming Zhang, Kailun Yang, Alina Roitberg 외

Intelligent vehicles clearly benefit from the expanded Field of View (FoV) of the 360-degree sensors, but the vast majority of available semantic segmentation training images are captured with pinhole cameras. In this wo…

Domain AdaptationSegmentationSemantic SegmentationUnsupervised Domain Adaptation

Multi-source Domain Adaptation for Panoramic Semantic Segmentation

2024-08-29 · Jing Jiang, Sicheng Zhao, Jiankun Zhu, Wenbo Tang 외

Unsupervised domain adaptation methods for panoramic semantic segmentation utilize real pinhole images or low-cost synthetic panoramic images to transfer segmentation models to real panoramic images. However, these metho…

Domain AdaptationSegmentationSemantic SegmentationUnsupervised Domain Adaptation

PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation

2023-09-21 · Shilin Yan, Xiaohao Xu, Renrui Zhang, Lingyi Hong 외

Panoramic videos contain richer spatial information and have attracted tremendous amounts of attention due to their exceptional experience in some fields such as autonomous driving and virtual reality. However, existing …

Autonomous DrivingSegmentationSemantic SegmentationVideo Object Segmentation+2