paper-with-me

홈 › Papers

BelHouse3D: A Benchmark Dataset for Assessing Occlusion Robustness in 3D Point Cloud Semantic Segmentation

2024-11-20 · Umamaheswaran Raman Kumar, Abdur Razzaq Fayjie, Jurgen Hannaert, Patrick Vandewalle

Large-scale 2D datasets have been instrumental in advancing machine learning; however, progress in 3D vision tasks has been relatively slow. This disparity is largely due to the limited availability of 3D benchmarking datasets. In particular, creating real-world point cloud datasets for indoor scene semantic segmentation presents considerable challenges, including data collection within confined spaces and the costly, often inaccurate process of per-point labeling to generate ground truths. While synthetic datasets address some of these challenges, they often fail to replicate real-world conditions, particularly the occlusions that occur in point clouds collected from real environments. Existing 3D benchmarking datasets typically evaluate deep learning models under the assumption that training and test data are independently and identically distributed (IID), which affects the models' usability for real-world point cloud segmentation. To address these challenges, we introduce the BelHouse3D dataset, a new synthetic point cloud dataset designed for 3D indoor scene semantic segmentation. This dataset is constructed using real-world references from 32 houses in Belgium, ensuring that the synthetic data closely aligns with real-world conditions. Additionally, we include a test set with data occlusion to simulate out-of-distribution (OOD) scenarios, reflecting the occlusions commonly encountered in real-world point clouds. We evaluate popular point-based semantic segmentation methods using our OOD setting and present a benchmark. We believe that BelHouse3D and its OOD setting will advance research in 3D point cloud semantic segmentation for indoor scenes, providing valuable insights for the development of more generalizable models.

📄 PDF Abstract BibTeX arXiv:2411.13251

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingPoint Cloud SegmentationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

I Know You Can't See Me: Dynamic Occlusion-Aware Safety Validation of Strategic Planners for Autonomous Vehicles Using Hypergames

2021-09-20 · Maximilian Kahn, Atrisha Sarkar, Krzysztof Czarnecki

A particular challenge for both autonomous and human driving is dealing with risk associated with dynamic occlusion, i.e., occlusion caused by other vehicles in traffic. Based on the theory of hypergames, we develop a no…

Autonomous Vehicles

OccRob: Efficient SMT-Based Occlusion Robustness Verification of Deep Neural Networks

2023-01-27 · Xingwu Guo, Ziwei Zhou, Yueling Zhang, Guy Katz 외

Occlusion is a prevalent and easily realizable semantic perturbation to deep neural networks (DNNs). It can fool a DNN into misclassifying an input image by occluding some segments, possibly resulting in severe errors. T…

How Robust is 3D Human Pose Estimation to Occlusion?

2018-08-28 · István Sárándi, Timm Linder, Kai O. Arras, Bastian Leibe

Occlusion is commonplace in realistic human-robot shared environments, yet its effects are not considered in standard 3D human pose estimation benchmarks. This leaves the question open: how robust are state-of-the-art 3D…

3D Human Pose Estimation3D Pose EstimationData AugmentationPose Estimation

Delving into High-Quality Synthetic Face Occlusion Segmentation Datasets

2022-05-12 · Kenny T. R. Voo, Liming Jiang, Chen Change Loy

This paper performs comprehensive analysis on datasets for occlusion-aware face segmentation, a task that is crucial for many downstream applications. The collection and annotation of such datasets are time-consuming and…

SegmentationSynthetic Data GenerationVocal Bursts Intensity Prediction

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

2026-05-13 · Haoren Zhao, Tianyi Chen, Zhen Wang arxiv

Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interfaces. This paper identifies a critical reliability gap: state-of-the…

Semantic Similarity