paper-with-me

Papers

Multi-modality Affinity Inference for Weakly Supervised 3D Semantic Segmentation

2023-12-27 · Xiawei Li, Qingyuan Xu, Jing Zhang, Tianyi Zhang, Qian Yu, Lu Sheng, Dong Xu

3D point cloud semantic segmentation has a wide range of applications. Recently, weakly supervised point cloud segmentation methods have been proposed, aiming to alleviate the expensive and laborious manual annotation process by leveraging scene-level labels. However, these methods have not effectively exploited the rich geometric information (such as shape and scale) and appearance information (such as color and texture) present in RGB-D scans. Furthermore, current approaches fail to fully leverage the point affinity that can be inferred from the feature extraction network, which is crucial for learning from weak scene-level labels. Additionally, previous work overlooks the detrimental effects of the long-tailed distribution of point cloud data in weakly supervised 3D semantic segmentation. To this end, this paper proposes a simple yet effective scene-level weakly supervised point cloud segmentation method with a newly introduced multi-modality point affinity inference module. The point affinity proposed in this paper is characterized by features from multiple modalities (e.g., point cloud and RGB), and is further refined by normalizing the classifier weights to alleviate the detrimental effects of long-tailed distribution without the need of the prior of category distribution. Extensive experiments on the ScanNet and S3DIS benchmarks verify the effectiveness of our proposed method, which outperforms the state-of-the-art by ~4% to ~6% mIoU. Codes are released at https://github.com/Sunny599/AAAI24-3DWSSG-MMA.

📄 PDF Abstract BibTeX arXiv:2312.16578

Code (1)

sunny599/aaai24-3dwssg-mma 공식 구현

Tasks

3D Semantic SegmentationPoint Cloud SegmentationSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

Affinity Mixup for Weakly Supervised Sound Event Detection

2021-06-21 · Mohammad Rasool Izadi, Robert Stevenson, Laura N. Kloepper

The weakly supervised sound event detection problem is the task of predicting the presence of sound events and their corresponding starting and ending points in a weakly labeled dataset. A weak dataset associates each tr…

Event DetectionSound Event Detection

Revisiting Weakly-Supervised Video Scene Graph Generation via Pair Affinity Learning

2026-03-23 · Minseok Kang, Minhyeok Lee, Minjung Kim, Jungho Lee 외 arxiv

Weakly-supervised video scene graph generation (WS-VSGG) aims to parse video content into structured relational triplets without bounding box annotations and with only sparse temporal labeling, significantly reducing ann…

Video scene graph generation

Tackling Ambiguity from Perspective of Uncertainty Inference and Affinity Diversification for Weakly Supervised Semantic Segmentation

2024-04-12 · Zhiwei Yang, Yucong Meng, Kexue Fu, Shuo Wang 외

Weakly supervised semantic segmentation (WSSS) with image-level labels intends to achieve dense tasks without laborious annotations. However, due to the ambiguous contexts and fuzzy regions, the performance of WSSS, espe…

DiversitySemantic SegmentationWeakly supervised Semantic SegmentationWeakly-Supervised Semantic Segmentation

Adaptive Affinity Loss and Erroneous Pseudo-Label Refinement for Weakly Supervised Semantic Segmentation

2021-08-03 · Xiangrong Zhang, Zelin Peng, Peng Zhu, Tianyang Zhang 외

Semantic segmentation has been continuously investigated in the last ten years, and majority of the established technologies are based on supervised models. In recent years, image-level weakly supervised semantic segment…

Pseudo LabelSegmentationSemantic SegmentationWeakly supervised Semantic Segmentation+1

Leveraging Auxiliary Tasks with Affinity Learning for Weakly Supervised Semantic Segmentation

2021-07-25 · ICCV 2021 10 · Lian Xu, Wanli Ouyang, Mohammed Bennamoun, Farid Boussaid 외

Semantic segmentation is a challenging task in the absence of densely labelled data. Only relying on class activation maps (CAM) with image-level labels provides deficient segmentation supervision. Prior works thus consi…

Auxiliary Learningimage-classificationImage ClassificationMulti-Label Image Classification+7