paper-with-me

홈 › Papers

POIRot: A rotation invariant omni-directional pointnet

2019-10-29 · Liu Yang, Rudrasis Chakraborty, Stella X. Yu

Point-cloud is an efficient way to represent 3D world. Analysis of point-cloud deals with understanding the underlying 3D geometric structure. But due to the lack of smooth topology, and hence the lack of neighborhood structure, standard correlation can not be directly applied on point-cloud. One of the popular approaches to do point correlation is to partition the point-cloud into voxels and extract features using standard 3D correlation. But this approach suffers from sparsity of point-cloud and hence results in multiple empty voxels. One possible solution to deal with this problem is to learn a MLP to map a point or its local neighborhood to a high dimensional feature space. All these methods suffer from a large number of parameters requirement and are susceptible to random rotations. A popular way to make the model "invariant" to rotations is to use data augmentation techniques with small rotations but the potential drawback includes \item more training samples \item susceptible to large rotations. In this work, we develop a rotation invariant point-cloud segmentation and classification scheme based on the omni-directional camera model (dubbed as {\bf POIRot$^1$}). Our proposed model is rotationally invariant and can preserve geometric shape of a 3D point-cloud. Because of the inherent rotation invariant property, our proposed framework requires fewer number of parameters (please see \cite{Iandola2017SqueezeNetAA} and the references therein for motivation of lean models). Several experiments have been performed to show that our proposed method can beat the state-of-the-art algorithms in classification and part segmentation applications.

📄 PDF Abstract BibTeX arXiv:1910.13050

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationPoint Cloud Segmentation

Similar Papers 제목 키워드 기반

Rotation-Invariant Autoencoders for Signals on Spheres

2020-12-08 · Suhas Lohit, Shubhendu Trivedi

Omnidirectional images and spherical representations of $3D$ shapes cannot be processed with conventional 2D convolutional neural networks (CNNs) as the unwrapping leads to large distortion. Using fast implementations of…

ClusteringRetrieval

Are Multimodal Large Language Models Ready for Omnidirectional Spatial Reasoning?

2025-05-17 · Zihao Dongfang, Xu Zheng, Ziqiao Weng, Yuanhuiyi Lyu 외

The 180x360 omnidirectional field of view captured by 360-degree cameras enables their use in a wide range of applications such as embodied AI and virtual reality. Although recent advances in multimodal large language mo…

HallucinationObject CountingSpatial Reasoning

Orientation-aware Semantic Segmentation on Icosahedron Spheres

2019-07-30 · ICCV 2019 10 · Chao Zhang, Stephan Liwicki, William Smith, Roberto Cipolla

We address semantic segmentation on omnidirectional images, to leverage a holistic understanding of the surrounding scene for applications like autonomous driving systems. For the spherical domain, several methods recent…

Autonomous DrivingSemantic Segmentation

Distortion-Tolerant Monocular Depth Estimation On Omnidirectional Images Using Dual-cubemap

2022-03-18 · Zhijie Shen, Chunyu Lin, Lang Nie, Kang Liao 외

Estimating the depth of omnidirectional images is more challenging than that of normal field-of-view (NFoV) images because the varying distortion can significantly twist an object's shape. The existing methods suffer fro…

Depth EstimationMonocular Depth Estimation

Omni-directional Pathloss Measurement Based on Virtual Antenna Array with Directional Antennas

2022-08-07 · Mengting Li, Fengchun Zhang, Xiang Zhang, Yejian Lyu 외

Omni-directional pathloss, which refers to the pathloss when omni-directional antennas are used at the link ends, are essential for system design and evaluation. In the millimeter-wave (mm-Wave) and beyond bands, high ga…