PS^2-Net: A Locally and Globally Aware Network for Point-Based Semantic Segmentation
In this paper, we present the PS^2-Net -- a locally and globally aware deep learning framework for semantic segmentation on 3D scene-level point clouds. In order to deeply incorporate local structures and global context to support 3D scene segmentation, our network is built on four repeatedly stacked encoders, where each encoder has two basic components: EdgeConv that captures local structures and NetVLAD that models global context. Different from existing start-of-the-art methods for point-based scene semantic segmentation that either violate or do not achieve permutation invariance, our PS^2-Net is designed to be permutation invariant which is an essential property of any deep network used to process unordered point clouds. We further provide theoretical proof to guarantee the permutation invariance property of our network. We perform extensive experiments on two large-scale 3D indoor scene datasets and demonstrate that our PS2-Net is able to achieve state-of-the-art performances as compared to existing approaches.
Code (1)
Tasks
Scene SegmentationSegmentationSemantic SegmentationSimilar Papers 제목 키워드 기반
Fusion-Aware Point Convolution for Online Semantic 3D Scene Segmentation
Online semantic 3D segmentation in company with real-time RGB-D reconstruction poses special challenges such as how to perform 3D convolution directly over the progressively fused 3D geometric data, and how to smartly fu…
RGB-D ReconstructionScene SegmentationSFD2: Semantic-guided Feature Detection and Description
Visual localization is a fundamental task for various applications including autonomous driving and robotics. Prior methods focus on extracting large amounts of often redundant locally reliable features, resulting in lim…
2k4kAutonomous DrivingVisual LocalizationGlobal-Local Medical SAM Adaptor Based on Full Adaption
Emerging of visual language models, such as the segment anything model (SAM), have made great breakthroughs in the field of universal semantic segmentation and significantly aid the improvements of medical image segmenta…
Image SegmentationMedical Image SegmentationSegmentationSemantic SegmentationLearning Multi-level Region Consistency with Dense Multi-label Networks for Semantic Segmentation
Semantic image segmentation is a fundamental task in image understanding. Per-pixel semantic labelling of an image benefits greatly from the ability to consider region consistency both locally and globally. However, many…
Image SegmentationSegmentationSemantic Segmentation7th AI Driving Olympics: 1st Place Report for Panoptic Tracking
In this technical report, we describe our EfficientLPT architecture that won the panoptic tracking challenge in the 7th AI Driving Olympics at NeurIPS 2021. Our architecture builds upon the top-down EfficientLPS panoptic…
BenchmarkingPanoptic SegmentationPanoptic Tracking