paper-with-me

홈 › Papers

Open Panoramic Segmentation

2024-07-02 · Junwei Zheng, Ruiping Liu, Yufan Chen, Kunyu Peng, Chengzhi Wu, Kailun Yang, Jiaming Zhang, Rainer Stiefelhagen

Panoramic images, capturing a 360{\deg} field of view (FoV), encompass omnidirectional spatial information crucial for scene understanding. However, it is not only costly to obtain training-sufficient dense-annotated panoramas but also application-restricted when training models in a close-vocabulary setting. To tackle this problem, in this work, we define a new task termed Open Panoramic Segmentation (OPS), where models are trained with FoV-restricted pinhole images in the source domain in an open-vocabulary setting while evaluated with FoV-open panoramic images in the target domain, enabling the zero-shot open panoramic semantic segmentation ability of models. Moreover, we propose a model named OOOPS with a Deformable Adapter Network (DAN), which significantly improves zero-shot panoramic semantic segmentation performance. To further enhance the distortion-aware modeling ability from the pinhole source domain, we propose a novel data augmentation method called Random Equirectangular Projection (RERP) which is specifically designed to address object deformations in advance. Surpassing other state-of-the-art open-vocabulary semantic segmentation approaches, a remarkable performance boost on three panoramic datasets, WildPASS, Stanford2D3D, and Matterport3D, proves the effectiveness of our proposed OOOPS model with RERP on the OPS task, especially +2.2% on outdoor WildPASS and +2.4% mIoU on indoor Stanford2D3D. The source code is publicly available at https://junweizheng93.github.io/publications/OPS/OPS.html.

📄 PDF Abstract BibTeX arXiv:2407.02685

Code (1)

JunweiZheng93/OPS 공식 구현 pytorch

Tasks

Open-Vocabulary Panoramic Semantic Segmentation

Methods 이 논문이 사용한 방법론

Adapter 설명 없음

Similar Papers 제목 키워드 기반

JOPP-3D: Joint Open Vocabulary Semantic Segmentation on Point Clouds and Panoramas

2026-03-06 · Sandeep Inuganti, Hideaki Kanayama, Kanta Shimizu, Mahdi Chamseddine 외 arxiv

Semantic segmentation across visual modalities such as 3D point clouds and panoramic images remains a challenging task, primarily due to the scarcity of annotated data and the limited adaptability of fixed-label models. …

Open Vocabulary Semantic Segmentation3D Semantic SegmentationScene UnderstandingPoint Clouds

Waymo Open Dataset: Panoramic Video Panoptic Segmentation

2022-06-15 · Jieru Mei, Alex Zihao Zhu, Xinchen Yan, Hang Yan 외

Panoptic image segmentation is the computer vision task of finding groups of pixels in an image and assigning semantic classes and object instance identifiers to them. Research in image segmentation has become increasing…

3D Multi-Object TrackingAutonomous DrivingAutonomous VehiclesImage Segmentation+6

Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation

2026-03-16 · Yuanfan Zheng, Kunyu Peng, Xu Zheng, Kailun Yang arxiv

Cross-domain panoramic semantic segmentation has attracted growing interest as it enables comprehensive 360° scene understanding for real-world applications. However, it remains particularly challenging due to severe geo…

Semantic SegmentationScene UnderstandingDomain AdaptationGraph Matching

Multi-source Domain Adaptation for Panoramic Semantic Segmentation

2024-08-29 · Jing Jiang, Sicheng Zhao, Jiankun Zhu, Wenbo Tang 외

Unsupervised domain adaptation methods for panoramic semantic segmentation utilize real pinhole images or low-cost synthetic panoramic images to transfer segmentation models to real panoramic images. However, these metho…

Domain AdaptationSegmentationSemantic SegmentationUnsupervised Domain Adaptation

PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation

2023-09-21 · Shilin Yan, Xiaohao Xu, Renrui Zhang, Lingyi Hong 외

Panoramic videos contain richer spatial information and have attracted tremendous amounts of attention due to their exceptional experience in some fields such as autonomous driving and virtual reality. However, existing …

Autonomous DrivingSegmentationSemantic SegmentationVideo Object Segmentation+2