paper-with-me

홈 › Papers

Capturing Omni-Range Context for Omnidirectional Segmentation

2021-03-09 · CVPR 2021 1 · Kailun Yang, Jiaming Zhang, Simon Reiß, Xinxin Hu, Rainer Stiefelhagen

Convolutional Networks (ConvNets) excel at semantic segmentation and have become a vital component for perception in autonomous driving. Enabling an all-encompassing view of street-scenes, omnidirectional cameras present themselves as a perfect fit in such systems. Most segmentation models for parsing urban environments operate on common, narrow Field of View (FoV) images. Transferring these models from the domain they were designed for to 360-degree perception, their performance drops dramatically, e.g., by an absolute 30.0% (mIoU) on established test-beds. To bridge the gap in terms of FoV and structural distribution between the imaging domains, we introduce Efficient Concurrent Attention Networks (ECANets), directly capturing the inherent long-range dependencies in omnidirectional imagery. In addition to the learned attention-based contextual priors that can stretch across 360-degree images, we upgrade model training by leveraging multi-source and omni-supervised learning, taking advantage of both: Densely labeled and unlabeled data originating from multiple datasets. To foster progress in panoramic image segmentation, we put forward and extensively evaluate models on Wild PAnoramic Semantic Segmentation (WildPASS), a dataset designed to capture diverse scenes from all around the globe. Our novel model, training regimen and multi-source prediction fusion elevate the performance (mIoU) to new state-of-the-art results on the public PASS (60.2%) and the fresh WildPASS (69.0%) benchmarks.

📄 PDF Abstract BibTeX arXiv:2103.05687

Code (1)

elnino9ykl/WildPASS 공식 구현 pytorch

Tasks

Autonomous DrivingImage SegmentationSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

360VOTS: Visual Object Tracking and Segmentation in Omnidirectional Videos

2024-04-22 · Yinzhe Xu, Huajian Huang, Yingshu Chen, Sai-Kit Yeung

Visual object tracking and segmentation in omnidirectional videos are challenging due to the wide field-of-view and large spherical distortion brought by 360{\deg} images. To alleviate these problems, we introduce a nove…

ObjectObject TrackingSegmentationSemantic Segmentation+3

Segmentation-Based Bounding Box Generation for Omnidirectional Pedestrian Detection

2021-04-28 · Masato Tamura, Tomoaki Yoshinaga

We propose a segmentation-based bounding box generation method for omnidirectional pedestrian detection that enables detectors to tightly fit bounding boxes to pedestrians without omnidirectional images for training. Due…

object-detectionObject DetectionPedestrian Detection

FreDSNet: Joint Monocular Depth and Semantic Segmentation with Fast Fourier Convolutions

2022-10-04 · Bruno Berenguel-Baeta, Jesus Bermudez-Cameo, Jose J. Guerrero

In this work we present FreDSNet, a deep learning solution which obtains semantic 3D understanding of indoor environments from single panoramas. Omnidirectional images reveal task-specific advantages when addressing scen…

Depth EstimationMonocular Depth EstimationScene UnderstandingSegmentation+1

SC-OmniGS: Self-Calibrating Omnidirectional Gaussian Splatting

2025-02-07 · Huajian Huang, Yingshu Chen, Longwei Li, Hui Cheng 외

360-degree cameras streamline data collection for radiance field 3D reconstruction by capturing comprehensive scene data. However, traditional radiance field methods do not address the specific challenges inherent to 360…

3D Reconstruction

Applications of Deep Learning for Top-View Omnidirectional Imaging: A Survey

2023-04-17 · Jingrui Yu, Ana Cecilia Perez Grassi, Gangolf Hirtz

A large field-of-view fisheye camera allows for capturing a large area with minimal numbers of cameras when they are mounted on a high position facing downwards. This top-view omnidirectional setup greatly reduces the wo…

Activity RecognitionDeep LearningMiscellaneousobject-detection+3